jobs
Safeguards Enforcement Analyst, User Well-being
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers,
In inglese
- Company
- Anthropic
- Location
- San Francisco, CA; New York City, NY; Washington, DC
- Status
- Open
- Posted
- 2026-08-03T18:48:40+00:00
Anthropic's User Well-being team is hiring a Safeguards Enforcement Analyst in San Francisco, New York City, or Washington, DC. The analyst designs and deploys mental-health guardrails: defining metrics, curating evaluation datasets, helping Engineering and Data Science tune detection models and thresholds, monitoring interventions, and reviewing flagged content. The posting asks for trust and safety, product policy or moderation experience involving suicide, self-harm, or related harms, plus SQL, rubric design, review-queue management, and prompt writing for generative AI. Preferred: mental-health subject expertise. It warns the role involves disturbing content.
Original job posting