jobs
Research Engineer, RL Engineering
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers,
- Company
- Anthropic
- Location
- San Francisco, CA; New York City, NY; Seattle, WA
- Status
- Open
- Posted
- 2025-10-15T00:44:37+00:00
Anthropic's Reinforcement Learning Engineering team is hiring an ML systems engineer to build the algorithms and systems its finetuning researchers use to train Claude with RLHF. Work includes profiling the RL pipeline, monitoring training jobs, adapting finetuning systems to new model architectures and diagnosing slowdowns. The posting asks for 4+ years of software engineering experience; distributed systems, large-scale LLM training, Python and RLHF implementation are listed as pluses. Roles sit in San Francisco, New York or Seattle, with staff expected in office at least 25% of the time.
Original job posting