← All roles

Research Engineer, Interpretability

Anthropic ·US ·Hybrid, Remote 59d ago
PythonRustGoJavaPyTorchCUDADistributed Systems

About the role

Anthropic's Interpretability team is seeking a Research Engineer to build specialized inference and training infrastructure for mechanistic interpretability research. The role involves instrumenting model internals, resolving scaling bottlenecks, designing tools for researchers, and bringing research into production safety audits. The work is akin to reverse-engineering neural networks to understand how they work, with a focus on AI safety and reliability.

Requirements

5-10+ years of software engineering experience, proficiency in Python and at least one other language (e.g., Rust, Go, Java), strong ability to prioritize and collaborate, curiosity about AI interpretability and safety, comfort working with researchers. Preferred: experience with large-scale distributed systems, LLM optimization, PyTorch/CUDA or JAX/XLA, and building research tooling.

About the company

Anthropic

Anthropic is an AI safety and research company developing advanced AI systems like Claude. Their value proposition lies in building beneficial AI for humanity by prioritizing safety and ethical considerations throughout their research and product development.

View company page →
$315k–560k
Hybrid, Remote San Francisco, CASan Francisco, United States
Posted 59d ago
Apply for this role
Opens job-boards.greenhouse.io ↗
About the company
Anthropic
https://www.anthropic.com

Anthropic is an AI safety and research company developing advanced AI systems like Claude. Their value proposition lies in building beneficial AI for humanity by prioritizing safety and ethical considerations throughout their research and product development.

View company page →