jobs
Safeguards Enforcement Analyst, Safety Evaluations
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers,
- Company
- Anthropic
- Location
- San Francisco, CA; New York City, NY; Washington, DC
- Status
- Open
- Posted
- 2026-03-09T16:55:38+00:00
Anthropic's Safeguards team is hiring a Safeguards Enforcement Analyst for safety evaluations. The work covers running and monitoring model evaluations, interpreting results, flagging regressions, coordinating new evals with policy experts, driving mitigations, and writing documentation and processes so the work scales. The posting asks for trust-and-safety, content operations or policy enforcement experience, zero-to-one process building, program management, and data tools such as SQL. It is based in San Francisco, New York City or Washington, DC. Listed salary: $230,000–$270,000.
Original job posting