Vai al contenuto
AI.info

jobs

Research Engineer, Code RL (Reinforcement Learning)

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers,

In inglese

Company
Anthropic
Location
San Francisco, CA; New York City, NY
Status
Open
Posted
2026-06-11T21:47:04+00:00

Anthropic is hiring a Research Engineer for its Code RL team, focused on improving Claude models' software-engineering abilities. The role combines RL research and engineering: designing coding tasks and RL environments, building reward signals and verifiers, running training experiments on frontier models, and diagnosing model performance. The posting asks for strong Python and async/concurrent programming, end-to-end system ownership, and rigorous experimental design. Distinctive aspects include work on agentic coding, long-horizon autonomous engineering, and high-performance accelerator code, with collaboration across alignment and production training teams.

Original job posting