Safeguards Enforcement Analyst, Conventional Weapons

Anthropic
New York, San Francisco, Washington
Workplace: OnsiteFull timeUSD 245,000 - 330,000 annuallyFunction: Administration & Executive AssistanceEducation: bachelorsSkills: ["Communication","Stakeholder management","Policy analysis","Risk assessment"]

Build and scale automated safeguards enforcement for conventional weapons and dangerous technology misuse. Design enforcement workflows, review flagged content to make enforcement decisions, and develop evals to measure model performance and surface regressions and policy gaps. Partner with Engineering and Data Science to improve detection systems, maintain enforcement guidelines and reviewer documentation, and track emerging weapons and misuse trends to inform ongoing policy enforcement best practices.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
3 days ago

Safeguards Enforcement Analyst, Conventional Weapons

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Build and scale automated safeguards enforcement for conventional weapons and dangerous technology misuse. Design enforcement workflows, review flagged content to make enforcement decisions, and develop evals to measure model performance and surface regressions and policy gaps. Partner with Engineering and Data Science to improve detection systems, maintain enforcement guidelines and reviewer documentation, and track emerging weapons and misuse trends to inform ongoing policy enforcement best practices.
Location: New York, San Francisco, Washington
Workplace: Onsite
Employment Type: Full time
Job Function: Administration & Executive Assistance

Key Responsibilities

  • •Design and architect automated enforcement systems and scalable review workflows while maintaining high accuracy.
  • •Develop and maintain evals to measure model performance on policy areas and surface regressions and gaps.
  • •Partner with Engineering and Data Science to optimize detection and automated enforcement systems for policy violations.
  • •Review flagged content to drive enforcement decisions, paying attention to novel, technically sophisticated misuse attempts and emerging tactics.
  • •Develop enforcement guidelines and reviewer documentation, and keep workflows updated with emerging weapons trends, regulatory changes, and AI policy enforcement best practices.

Pay and Benefits

Salary: USD 245,000 - 330,000 annually
Perks:Paid LeaveParental LeaveEquity

Key Requirements

  • •Deep, applied expertise in weapons systems and the ability to translate technical evidence into enforcement decisions.
  • •Experience in policy enforcement, threat intelligence, counterterrorism, government, or a closely related field with direct exposure to harmful content or dangerous technology.
  • •Experience standing up and scaling policy enforcement or content review workflows.
  • •Proficiency in SQL and/or other data analysis tools to draw insights from large datasets and monitor enforcement workflow health.
  • •Experience developing prompts for content review and enforcement in the context of generative AI products.
Experience:AI safetyThreat intelligenceCounterterrorismContent moderationGenerative AI
Education:Bachelor's
Skills:CommunicationStakeholder managementPolicy analysisRisk assessment
Languages:English
Tech Stack:SQLPythonLarge language modelsGenerative AIOSINTMITRE ATT&CK

Eligibility

Work Authorization:Sponsorship available.

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn