Safeguards Enforcement Analyst, Chem & Explosives Harms

Anthropic
San Francisco, New York, Washington
Workplace: HybridFull timeUSD 245,000 - 285,000 annuallyFunction: CybersecurityEducation: bachelorsSkills: ["Clear communication","Decision-making","Proactiveness","Self-direction","Ability to handle ambiguity"]

Enforce a Usage Policy focused on chemical and explosives harms by monitoring platform activity, investigating potential violations, and escalating credible risks. Build and continuously improve end-to-end enforcement workflows and automated detection systems, ensuring accurate triage across complex content. Partner with Policy, Threat Intelligence, Data Science, and Engineering to analyze emerging patterns, optimize detection models, and provide enforcement-grounded feedback that strengthens safeguards at scale.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
1 month ago

Safeguards Enforcement Analyst, Chem & Explosives Harms

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 10 hours agoStatus: Live

Job Summary

Enforce a Usage Policy focused on chemical and explosives harms by monitoring platform activity, investigating potential violations, and escalating credible risks. Build and continuously improve end-to-end enforcement workflows and automated detection systems, ensuring accurate triage across complex content. Partner with Policy, Threat Intelligence, Data Science, and Engineering to analyze emerging patterns, optimize detection models, and provide enforcement-grounded feedback that strengthens safeguards at scale.
Location: San Francisco, New York, Washington
Workplace: Hybrid
Employment Type: Full time
Job Function: Cybersecurity

Key Responsibilities

  • •Enforce Usage Policies focused on detecting and mitigating chemical and explosives risks and harmful use of AI systems.
  • •Own and improve enforcement monitoring workflows, including end-to-end detection, investigation, triage, and escalation processes.
  • •Monitor and analyze platform activity to identify emerging chemical and explosives threats that may require policy updates or enforcement action.
  • •Design and architect automated enforcement systems and review workflows to scale while maintaining high accuracy.
  • •Conduct thorough investigations and partner across Policy, Threat Intelligence, Data Science, and Engineering to optimize detection models and strengthen safeguards.

Pay and Benefits

Salary: USD 245,000 - 285,000 annually
Perks:Paid LeaveParental Leave

Key Requirements

  • •Hold a degree in a chemistry-related field (e.g., chemistry, chemical engineering, materials science) and/or relevant professional experience.
  • •Have experience in Trust & Safety, content moderation, or policy enforcement at platform scale, including working with generative AI tools for review and enforcement workflows.
  • •Use AI tools to develop data dashboards for metrics collection and continuous improvement.
  • •Analyze complex, ambiguous situations and make well-reasoned, defensible decisions under time pressure.
  • •Communicate clearly in writing and translate technical chemistry concepts for diverse technical and non-technical audiences.
Experience:Trust & safetyContent moderationGenerative AI
Education:Bachelor's in Chemistry-related field
Skills:Clear communicationDecision-makingProactivenessSelf-directionAbility to handle ambiguity
Languages:English
Tech Stack:SQL

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn