Safeguards Enforcement Analyst, Cyber Harm
Anthropic
San Francisco, New York, Washington
Workplace: HybridFull timeUSD 285,000 - 330,000 annuallyFunction: CybersecurityEducation: bachelorsSkills: ["Communication","Stakeholder collaboration","Risk identification"]Review flagged content and accounts to make accurate enforcement decisions focused on detecting and mitigating attempts to misuse AI systems for malicious cyber operations. Triage ambiguous and high-severity cases, escalate when needed, and provide detailed feedback to the Safeguards policy team. Partner with Engineering and Data Science by surfacing detection model errors and quality signals to improve precision and recall, while staying current on evolving cyber threats and AI enforcement best practices.

