Vai al contenuto
AI.info

jobs

Research Engineer, RL Scaling Science

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers,

In inglese

Company
Anthropic
Location
London, UK
Status
Open
Posted
2026-06-22T11:06:02+00:00

Anthropic's RL Scaling Science team studies how reinforcement learning behaves as model size, compute and task horizon grow, and turns that into training recipes for frontier models. A research engineer designs and runs large-scale RL experiments, builds benchmarks that make long-horizon progress measurable, and ships validated findings into production training, debugging failures that appear only at scale. The posting asks for empirical RL or large-scale ML training skills, Python and distributed systems experience, and comfort at the research-systems boundary. London-based, hybrid.

Original job posting