Product Manager, Safeguards (Account Integrity & Abuse)

Anthropic
San Francisco, New York
Workplace: OnsiteFull timeUSD 385,000 - 460,000 annuallyFunction: Product ManagementEducation: bachelorsSkills: ["Ruthless prioritization","Cross-functional collaboration","Judgment","Written communication","Verbal communication"]

Own the ideation, design, development, and deployment of Safeguards systems to protect users and mitigate ethical, technical, and social risks from generative AI. Partner with research, policy, enforcement, and engineering to build detections, evals, interventions, and tools, defining tradeoffs and requirements from MVP to ideal state. Advance safety metrics and communicate effectively about safety across product UX surfaces like Claude.ai and the 1P API.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
2 days ago

Product Manager, Safeguards (Account Integrity & Abuse)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Own the ideation, design, development, and deployment of Safeguards systems to protect users and mitigate ethical, technical, and social risks from generative AI. Partner with research, policy, enforcement, and engineering to build detections, evals, interventions, and tools, defining tradeoffs and requirements from MVP to ideal state. Advance safety metrics and communicate effectively about safety across product UX surfaces like Claude.ai and the 1P API.
Location: San Francisco, New York
Workplace: Onsite
Employment Type: Full time
Job Function: Product Management
Seniority: Mid level

Key Responsibilities

  • •Determine how to build safety by design upstream and leverage downstream defenses across frontier models and Anthropic products for users on different surfaces.
  • •Write safety evals and communicate externally about safety.
  • •Drive impact through clear problem definition, solution options, and tradeoffs to set requirements from MVP to ideal state.
  • •Align and collaborate with policy, enforcement, research, engineering, and other cross-functional stakeholders.
  • •Develop metrics to measure performance, blindspots, and inform future project planning.

Pay and Benefits

Salary: USD 385,000 - 460,000 annually
Perks:Parental LeavePaid Leave

Key Requirements

  • •Ability to make technical tradeoff decisions while working across policy experts, AI/ML research engineers, and software engineering teams to build safety systems.
  • •Strong understanding of how products are used, including Safeguards concerns, and how to deliver effective solutions.
  • •Demonstrated ability to develop product and engineering strategy across multiple cross-functional teams in a rapidly changing space.
  • •Experience designing and building metrics to evaluate risks, system performance, user impact, and to make clear tradeoffs.
  • •Ability to plan, build, launch, and measure new products/systems in a zero-to-one environment.
Experience:Product management
Education:Bachelor's
Skills:Ruthless prioritizationCross-functional collaborationJudgmentWritten communicationVerbal communication
Tech Stack:Claude.ai1P APICloud platformsAI safetyAI/MLMetricsEvalsDetectionsInterventionsDeployment risk mitigation

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn