Head of Policy Design, Societal Harms

Anthropic
San Francisco
Workplace: OnsiteFull timeUSD 330,000 - 395,000 annuallyFunction: Design (Product/UX/UI/Visual)Education: bachelorsSkills: ["Cross-team collaboration","Judgment","Escalation","Communication","Prioritization"]

Lead Anthropic’s policy design team within Safeguards (Trust & Safety), responsible for consumer harm boundaries across child safety, user well-being, harmful manipulation, and election integrity. Manage leaders who define and evolve policies, detection/enforcement systems, and product interventions. Coordinate decisions across the portfolio, set mitigation strategy aligned with model training and deployment, and serve as escalation for high-severity, ambiguous harms while partnering with research, product, engineering, and external experts.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
3 days ago

Head of Policy Design, Societal Harms

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 19 hours agoStatus: Live

Job Summary

Lead Anthropic’s policy design team within Safeguards (Trust & Safety), responsible for consumer harm boundaries across child safety, user well-being, harmful manipulation, and election integrity. Manage leaders who define and evolve policies, detection/enforcement systems, and product interventions. Coordinate decisions across the portfolio, set mitigation strategy aligned with model training and deployment, and serve as escalation for high-severity, ambiguous harms while partnering with research, product, engineering, and external experts.
Location: San Francisco
Workplace: Onsite
Employment Type: Full time
Job Function: Design (Product/UX/UI/Visual)
Seniority: Sr. Director level

Key Responsibilities

  • •Lead, develop, and grow managers and teams responsible for the consumer harms portfolio, including child safety, user well-being, harmful manipulation, and election integrity.
  • •Coordinate policy decisions across the portfolio and ensure they remain consistent, tracked, and clearly owned across harm areas and product surfaces.
  • •Set strategy for model-top mitigations—policies, detection/enforcement systems, and product interventions—aligned with the alignment training team.
  • •Prioritize competing harm areas for shared resources and make tradeoffs and rationale clear to leadership.
  • •Partner with engineering, data science, product, legal, and research so consumer harms considerations are represented from training through launch across deployed surfaces.

Pay and Benefits

Salary: USD 330,000 - 395,000 annually
Perks:Paid LeaveParental Leave

Key Requirements

  • •Experience leading teams, including managing managers or senior specialists, in AI safety, product policy, or a related field.
  • •Deep, applied familiarity with consumer harm areas (e.g., child safety, mental health/well-being, manipulation, election integrity) and strong judgment about differences in mechanism, severity, and mitigation.
  • •A track record of exceptional cross-team collaboration to reach shared decisions with teams you don’t control.
  • •Working understanding of how frontier models are developed and deployed, including training/fine-tuning cycles, evaluations, and launch processes, and how environments change risk and mitigations.
  • •Experience translating policy positions into enforceable, measurable mechanisms and communicating reasoning to both technical and non-technical audiences.
Experience:AI safetyProduct policyConsumer harmsGenerative AILLM evaluations
Education:Bachelor's
Skills:Cross-team collaborationJudgmentEscalationCommunicationPrioritization

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn