jobs
Research Engineer, Performance RL (Reinforcement Learning)
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers,
- Company
- Anthropic
- Location
- San Francisco, CA
- Status
- Open
- Posted
- 2026-03-23T16:27:59+00:00
Anthropic's Code RL team, part of its Reinforcement Learning organization, is hiring a Research Engineer to improve how models write correct, fast code for accelerators. The person designs and builds RL environments and evaluations, runs experiments, shapes the research roadmap, and delivers work into training runs, working with researchers, engineers and performance specialists. The posting asks for accelerator expertise (CUDA, ROCm, Triton, Pallas), JAX or PyTorch, and experience across kernels, model code and distributed systems. Reinforcement learning experience is preferred, not required.
Original job posting