jobs
Safeguards Enforcement Analyst, Violence & Extremism
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers,
In inglese
- Company
- Anthropic
- Location
- San Francisco, CA; New York City, NY; Washington, DC
- Status
- Open
- Posted
- 2026-07-15T00:47:43+00:00
Anthropic is hiring a Safeguards Enforcement Analyst for Violence & Extremism in San Francisco, New York City or Washington, DC. The work: designing automated enforcement systems and review workflows, building evals of model behavior on weapons, violent extremism and threats of violence, reviewing flagged content, and escalating novel misuse patterns. The posting asks for enforcement, threat-intelligence or counterterrorism experience, SQL skills and hands-on use of generative AI. Analysts should expect exposure to violent, graphic and hateful material. The office policy is hybrid, with at least 25% of time in office.
Original job posting