Product Manager, Safeguards (Cyber)

Anthropic
San Francisco, California
Workplace: HybridFull timeUSD 305,000 - 385,000 annuallyFunction: Product ManagementExperience: 5+ yearsEducation: bachelorsSkills: ["Communication","Ruthless prioritization","Technical judgment","Cross-functional collaboration","Strategic thinking"]

Own ideation through deployment for Safeguards systems that keep Anthropic’s frontier models helpful, harmless, and honest. Partner with research and product teams to build detections, evals, interventions, and tools that measure and mitigate deployment and user risk across surfaces like Claude.ai and 1P API. Define safety-by-design priorities, develop metrics and blindspot insights, and translate complex technical concepts for cross-functional and external audiences.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
7 months ago

Product Manager, Safeguards (Cyber)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 9 hours agoStatus: Live

Job Summary

Own ideation through deployment for Safeguards systems that keep Anthropic’s frontier models helpful, harmless, and honest. Partner with research and product teams to build detections, evals, interventions, and tools that measure and mitigate deployment and user risk across surfaces like Claude.ai and 1P API. Define safety-by-design priorities, develop metrics and blindspot insights, and translate complex technical concepts for cross-functional and external audiences.
Location: San Francisco, California
Workplace: Hybrid
Employment Type: Full time
Job Function: Product Management
Seniority: Mid level

Key Responsibilities

  • •Own the ideation, design, development, and deployment of Safeguards systems and relevant product UX for safe advancement of frontier models.
  • •Develop detections, evals, interventions, and tools to measure and mitigate deployment and user risks.
  • •Determine how to build safety by design upstream and leverage downstream defenses across models, AI products, and customer surfaces.
  • •Set clear requirements by defining problems, solution options, and business/technical tradeoffs for MVP vs. ideal state.
  • •Align and collaborate with policy, enforcement, research, engineering, and cross-functional stakeholders to plan risk mitigation and future work.

Pay and Benefits

Salary: USD 305,000 - 385,000 annually
Perks:Paid LeaveParental Leave

Key Requirements

  • •Ability to make technical tradeoff decisions and work across policy experts, AI/ML research, and software engineering teams to build safety systems.
  • •Strong user understanding of product usage, Safeguards concerns, and how to deliver effective solutions.
  • •Demonstrated experience building product and engineering strategy across cross-functional teams in a rapidly changing space.
  • •Experience designing metrics to evaluate risks, system performance, and user impact while making clear tradeoffs.
  • •Ability to navigate ambiguity—prioritize amid changing product specs and demonstrate judgment in zero-to-one planning, building, launching, and measuring systems.
Experience:5+ yearsAI safetyGenerative AIProduct management
Education:Bachelor's in A field relevant to the role
Skills:CommunicationRuthless prioritizationTechnical judgmentCross-functional collaborationStrategic thinking
Languages:English
Tech Stack:Claude.ai1P APICloud platformsAI evals

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn