jobs
Staff+ Software Engineer, Safeguards
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers,
- Company
- Anthropic
- Location
- San Francisco, CA; New York City, NY
- Status
- Open
- Posted
- 2025-10-14T18:45:18+00:00
The Safeguards team at Anthropic is hiring staff-level software engineers to build systems that monitor models, detect misuse and decide what happens when a safety system fires. Work spans agentic tooling for trust and safety, user-facing interventions across products and the API, and investigation infrastructure running across three clouds. The posting asks for a computer science degree or comparable experience, Python and TypeScript, and cross-stack ability; eight-plus years and abuse or fraud detection experience are listed as strengths.
Original job posting