Research Engineer – Interpretability
Anthropic · San Francisco
Job description
About the role
Anthropic is looking for a Research Engineer to join its Interpretability team. You will help reverse‑engineer large language models to build a mechanistic understanding that underpins safe and trustworthy AI. The work blends cutting‑edge research with production‑grade engineering to scale interpretability tools for real‑world safety audits.
Key responsibilities
- Build and maintain specialized inference and training infrastructure for interpretability research, including instrumented forward/backward passes, activation extraction, and steering‑vector application.
- Identify and resolve scaling and efficiency bottlenecks through profiling, optimization, and close collaboration with peer infrastructure teams.
- Design abstractions, tools, and platforms that enable researchers to experiment rapidly without hitting engineering barriers.
- Help transition interpretability research findings into production safety audits and related pipelines.
Required profile
- Experience building and operating large‑scale model training and inference infrastructure.
- Strong ability to profile systems, diagnose performance issues, and implement optimizations.
- Collaborative mindset with a track record of working closely with infrastructure and research teams.
- Interest in AI safety and interpretability research.
Required skills
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in the United States.
Salaries by job title
Apply in 30 seconds
Enter your email to apply. An account will be created automatically.
By continuing, you accept our terms of use.
Already have an account? Login
Published 4 weeks ago
Expires 1 month from now
12 views · 0 interested
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
Anthropic
San Francisco
Related job offers
-
Machine Learning Engineer
sift San Francisco -
Software Engineer – Horizon
sierra San Francisco -
Software Engineer – AI Agent for Tech, Media & Telecom
sierra San Francisco -
Director
Food Safety and Inspection Service Thitani location -
AI Safety Specialist (Short‑Term Project)
HumanitApp Thitani location