jobs
Research Engineer, Universes
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers,
- Company
- Anthropic
- Location
- Remote-Friendly (Travel-Required); San Francisco, CA; Seattle, WA; New York City, NY
- Status
- Open
- Posted
- 2026-01-09T16:41:32+00:00
Anthropic's Universes team, part of Research, is hiring research engineers to design realistic training environments where models practise long-horizon agentic work — ambiguity, interruptions, extended context, open-ended judgment — plus evaluations that measure real capability. The role mixes reinforcement learning research with production engineering: building environments, shipping them into training runs, and debugging across research and production ML stacks. The posting asks for strong software engineering, high agency, comfort with uncertainty and sound research judgment. Useful backgrounds include LLM training, fine-tuning or evaluation, RL environments or simulation, sandboxing and large-scale ML infrastructure. Pay is $500,000–$850,000.
Original job posting