Product Manager, Safeguards (Generalist)

Anthropic
San Francisco, New York
Workplace: OnsiteFull timeUSD 385,000 - 460,000 annuallyFunction: Product ManagementExperience: 5+ yearsEducation: bachelorsSkills: ["Communication","Collaboration","Prioritization","Judgment","Creative thinking"]

Own the ideation, design, development, and deployment of Safeguards systems that help protect users from misuse and ethical, technical, and social risks across Anthropic’s frontier-model products. Partner with research, policy, enforcement, and engineering to create detections, evals, interventions, and measurement tools. Define problems and requirements, prioritize ruthlessly, and develop metrics to track performance, blindspots, and user impact across cloud platforms and product surfaces like Claude.ai and the 1P API.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
11 hours ago

Product Manager, Safeguards (Generalist)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Own the ideation, design, development, and deployment of Safeguards systems that help protect users from misuse and ethical, technical, and social risks across Anthropic’s frontier-model products. Partner with research, policy, enforcement, and engineering to create detections, evals, interventions, and measurement tools. Define problems and requirements, prioritize ruthlessly, and develop metrics to track performance, blindspots, and user impact across cloud platforms and product surfaces like Claude.ai and the 1P API.
Location: San Francisco, New York
Workplace: Onsite
Employment Type: Full time
Job Function: Product Management
Seniority: Mid level

Key Responsibilities

  • •Determine how to embed safety by design upstream and apply downstream defenses across Anthropic’s frontier models and product surfaces (Claude.ai, 1P API, and external cloud providers).
  • •Write safety evals and communicate externally about safety.
  • •Drive impact through ruthless prioritization by defining problems, solution options, requirements, and MVP vs. ideal-state tradeoffs.
  • •Align and collaborate with policy, enforcement, research, engineering, and other cross-functional stakeholders.
  • •Lead development of metrics to measure performance, blindspots, deployment risks, and user impact to inform future planning.

Pay and Benefits

Salary: USD 385,000 - 460,000 annually
Perks:Paid LeaveParental LeaveFlexible Hours

Key Requirements

  • •Make technical tradeoff decisions and work across policy experts, AI/ML research engineers, and software engineering teams to design and build safety systems.
  • •Strong understanding of user workflows, Safeguards concerns, and how the product provides the best solutions.
  • •Demonstrate ability to build product and engineering strategy across multiple cross-functional teams in a rapidly changing space.
  • •Design and build metrics to evaluate risk, system performance, user impact, and make clear tradeoffs.
  • •Plan, build, launch, and measure new products/systems in a zero-to-one environment, with strong written and verbal communication.
  • •Ability to clearly articulate complex technical concepts to non-technical audiences in written and verbal communication.
Experience:5+ years
Education:Bachelor's
Skills:CommunicationCollaborationPrioritizationJudgmentCreative thinking
Languages:English
Tech Stack:Claude.ai1P APICloud platforms

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn