Vai al contenuto
AI.info

jobs

Safeguards Enforcement Analyst, User Well-being

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers,

In inglese

Company
Anthropic
Location
San Francisco, CA; New York City, NY; Washington, DC
Status
Open
Posted
2026-08-03T18:48:40+00:00

Anthropic's User Well-being team is hiring a Safeguards Enforcement Analyst in San Francisco, New York City, or Washington, DC. The analyst designs and deploys mental-health guardrails: defining metrics, curating evaluation datasets, helping Engineering and Data Science tune detection models and thresholds, monitoring interventions, and reviewing flagged content. The posting asks for trust and safety, product policy or moderation experience involving suicide, self-harm, or related harms, plus SQL, rubric design, review-queue management, and prompt writing for generative AI. Preferred: mental-health subject expertise. It warns the role involves disturbing content.

Original job posting