Skip to content
AI.info

jobs

Research Engineer, Universes

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers,

Company
Anthropic
Location
Remote-Friendly (Travel-Required); San Francisco, CA; Seattle, WA; New York City, NY
Status
Open
Posted
2026-01-09T16:41:32+00:00

Anthropic's Universes team, part of Research, is hiring research engineers to design realistic training environments where models practise long-horizon agentic work — ambiguity, interruptions, extended context, open-ended judgment — plus evaluations that measure real capability. The role mixes reinforcement learning research with production engineering: building environments, shipping them into training runs, and debugging across research and production ML stacks. The posting asks for strong software engineering, high agency, comfort with uncertainty and sound research judgment. Useful backgrounds include LLM training, fine-tuning or evaluation, RL environments or simulation, sandboxing and large-scale ML infrastructure. Pay is $500,000–$850,000.

Original job posting