Safeguards Enforcement Lead, User Well-Being

Anthropic
New York, San Francisco, Washington DC
Workplace: OnsiteFull timeUSD 285,000 - 330,000 annuallyFunction: Administration & Executive AssistanceEducation: bachelorsSkills: ["Communication","Quality assurance","Stakeholder communication","Cross-functional collaboration","Trend identification"]

Lead enforcement operations for user well-being safeguards, managing workflows for child safety, mental health, abuse/exploitation, and age assurance. Own day-to-day content review partner processes (onboarding, training, QA, escalation), and design scalable enforcement workflows that maintain accuracy as volume grows. Partner with Engineering and Data Science to improve detection and automated enforcement systems, document decision guidelines, monitor trends, and coordinate required reporting (e.g., NCMEC) as laws and risks evolve.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
3 days ago

Safeguards Enforcement Lead, User Well-Being

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 5 hours agoStatus: Live

Job Summary

Lead enforcement operations for user well-being safeguards, managing workflows for child safety, mental health, abuse/exploitation, and age assurance. Own day-to-day content review partner processes (onboarding, training, QA, escalation), and design scalable enforcement workflows that maintain accuracy as volume grows. Partner with Engineering and Data Science to improve detection and automated enforcement systems, document decision guidelines, monitor trends, and coordinate required reporting (e.g., NCMEC) as laws and risks evolve.
Location: New York, San Francisco, Washington DC
Workplace: Onsite
Employment Type: Full time
Job Function: Administration & Executive Assistance
Seniority: Manager level

Key Responsibilities

  • •Manage a team of individual contributors across multiple policy areas under the User Well-Being banner.
  • •Serve as the primary point of contact for review partners, including onboarding, training, quality assurance, and relationship management.
  • •Design and improve enforcement workflows to scale as volume grows while maintaining accuracy and consistency.
  • •Partner with Engineering and Data Science teams to optimize detection models and automated enforcement systems.
  • •Coordinate reporting obligations to external bodies (e.g., NCMEC) and keep workflows aligned with evolving policies and legal frameworks.

Pay and Benefits

Salary: USD 285,000 - 330,000 annually
Perks:Paid LeaveParental LeaveEquity

Key Requirements

  • •Experience managing teams in the User Well-Being space.
  • •Experience in trust & safety, content moderation operations, or policy enforcement focused on child safety, mental health, abuse and exploitation, and age assurance.
  • •Experience managing or coordinating content review operations, including quality assurance and workflow management.
  • •Experience standing up and scaling policy enforcement or content review workflows.
  • •Proficiency in SQL and/or other data analysis tools to monitor workflow health and review queue metrics.
Education:Bachelor's
Skills:CommunicationQuality assuranceStakeholder communicationCross-functional collaborationTrend identification
Languages:English
Tech Stack:SQLPythonPhotoDNACSAI MatchHash-matchingPerceptual hashing

Eligibility

Work Authorization:Sponsorship available.

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn