Product Manager, Safeguards Rare Harms

Anthropic
San Francisco
Workplace: OnsiteFull timeUSD 305,000 - 385,000 annuallyFunction: Product ManagementExperience: 5+ yearsEducation: bachelorsSkills: ["Communication","Leadership","Problem-solving","Cross-functional collaboration","Strategic thinking"]

Product Manager for the Safeguards team at Anthropic, owning the ideation, design, development and deployment of Safeguards systems and UX to ensure frontier AI models are used safely across Claude.ai, 1P API, and external cloud surfaces. You’ll work with research, policy, enforcement, and engineering to develop detections, evals, interventions, and measurement tools to mitigate deployment and user risks in a fast-moving, ambiguous environment.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
3 months ago

Product Manager, Safeguards Rare Harms

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 3 hours agoStatus: Live

Job Summary

Product Manager for the Safeguards team at Anthropic, owning the ideation, design, development and deployment of Safeguards systems and UX to ensure frontier AI models are used safely across Claude.ai, 1P API, and external cloud surfaces. You’ll work with research, policy, enforcement, and engineering to develop detections, evals, interventions, and measurement tools to mitigate deployment and user risks in a fast-moving, ambiguous environment.
Location: San Francisco
Workplace: Onsite
Employment Type: Full time
Job Function: Product Management
Seniority: Sr. Manager level

Key Responsibilities

  • •Determine safety-by-design upstream and leverage downstream defenses for frontier models and products across multiple surfaces.
  • •Write safety evals and communicate externally about safety.
  • •Drive impact with ruthless prioritization by clearly defining problems, options, and MVP vs. ideal state with clear business and technical tradeoffs.
  • •Align and collaborate with policy, enforcement, research, engineering and cross-functional stakeholders.
  • •Lead the development of metrics to understand area performance, blindspots, and inform future project planning.

Pay and Benefits

Salary: USD 305,000 - 385,000 annually
Equity and Bonus:Equity
Perks:Paid LeaveParental LeaveFlexible HoursEquity

Key Requirements

  • •5+ years of product management experience with a track record of building roadmaps, analyzing data, and delivering measurable impact in fast-moving environments.
  • •Experience collaborating with policy, AI/ML research, software engineering, and cross-functional teams to design and build safety systems.
  • •Ability to understand user needs and Safeguards concerns, and translate them into product and engineering strategy.
  • •Proven ability to design and measure product metrics, risks, and tradeoffs to inform roadmap decisions.
  • •Bachelor’s degree or equivalent combination of education, training, and/or experience.
Experience:5+ yearsAI safetyMachine learningSaaSCloud
Education:Bachelor's
Skills:CommunicationLeadershipProblem-solvingCross-functional collaborationStrategic thinking
Languages:English
Tech Stack:Claude.ai1P APICloud platformsData metricsAPIs

Eligibility

Work Authorization:Sponsorship available.

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn