Research Engineer – AI Alignment
Anthropic · San Francisco
Job description
About the role
Anthropic is seeking a Research Engineer / Scientist to advance AI alignment research. You will design and execute rigorous machine‑learning experiments aimed at understanding and steering the behavior of powerful AI systems, ensuring they remain helpful, honest, and harmless.
Key responsibilities
- Build and run elegant, thorough ML experiments focused on AI safety and alignment.
- Collaborate with teams in Interpretability, Fine‑Tuning, and Frontier Red Team to test safety techniques.
- Develop and evaluate scalable oversight, AI control, and alignment stress‑testing methods.
- Conduct multi‑agent reinforcement‑learning experiments to probe model robustness.
- Investigate model welfare, moral status, and related safety assessments.
Required profile
- Passion for creating AI that is helpful, honest, and harmless.
- Interest in challenges posed by human‑level AI capabilities.
- Comfortable working as both a scientist and an engineer.
- Willingness to interview in Python and preferably reside in the Bay Area.
Required skills
- Python programming.
- Machine‑learning experimentation.
- Reinforcement learning, including multi‑agent setups.
- Experience with large language models and AI safety research.
What we offer
- Opportunity to work on cutting‑edge AI alignment problems.
- Collaboration with a fast‑growing, interdisciplinary research team.
- Access to state‑of‑the‑art AI resources and infrastructure.
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in the United States.
Salaries by job title
Apply in 30 seconds
Enter your email to apply. An account will be created automatically.
By continuing, you accept our terms of use.
Already have an account? Login
Published 4 weeks ago
Expires 1 month from now
14 views · 0 interested
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
Anthropic
San Francisco
Related job offers
-
Machine Learning Engineer
sift San Francisco -
Software Engineer – Horizon
sierra San Francisco -
Software Engineer – AI Agent for Tech, Media & Telecom
sierra San Francisco -
Director
Food Safety and Inspection Service Thitani location -
AI Safety Specialist (Short‑Term Project)
HumanitApp Thitani location