Software Engineer – Reinforcement Learning Data
Anthropic · San Francisco
Job description
About the role
Anthropic is looking for a senior Software Engineer to join its new RL Data team. You will shape the architecture of critical data pipelines, work closely with research groups, and ensure the quality and safety of reinforcement‑learning data that powers Claude.
Key responsibilities
- Own end‑to‑end components of the RL data stack, from architecture to operational maintenance.
- Design and implement data‑collection pipelines, iterate on prompts, evaluations, and grading mechanisms.
- Develop QA frameworks that detect reward‑hacking and guarantee environment integrity.
- Create user‑friendly interfaces for rapid human‑feedback collection.
- Harden execution environments with sandboxing, snapshotting, and comprehensive tool coverage.
- Collaborate with domain experts, operations, security, and compliance teams to roll out systems to new users and vendors.
Required profile
- Proven track record of leading major projects end‑to‑end in fast‑paced, ambiguous settings (e.g., founder, CTO, tech lead, or major open‑source contributor).
- Demonstrated ability to inspire teams, plan workstreams, and proactively resolve blockers.
- Strong software engineering expertise with a modern programming language.
- Experience using AI tools in daily workflows.
Required skills
- Python
- TypeScript
- Docker
- Kubernetes
- Common cloud infrastructure (e.g., AWS, GCP, Azure)
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in the United States.
Salaries by job title
Apply in 30 seconds
Enter your email to apply. An account will be created automatically.
By continuing, you accept our terms of use.
Already have an account? Login
Published 3 weeks ago
Expires 1 month from now
8 views · 0 interested
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
Anthropic
San Francisco