Staff+ Software Engineer, Safeguards Infrastructure

Anthropic
London
Workplace: OnsiteFull timeGBP 255,000 - 325,000 annuallyFunction: Software EngineeringExperience: 4-10 yearsEducation: bachelorsSkills: ["Communication","Teamwork","Problem-solving"]

Join Anthropic's Safeguards team to build foundational infrastructure for safety, oversight and intervention in AI systems. You will design systems to monitor models, detect unwanted behaviors, prevent misuse, and reduce human intervention. Work across the stack to support real-time safety defenses at scale, collaborating with researchers and policy experts to uphold safety, transparency, and responsible AI.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
7 months ago

Staff+ Software Engineer, Safeguards Infrastructure

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 20 hours agoStatus: Live

Job Summary

Join Anthropic's Safeguards team to build foundational infrastructure for safety, oversight and intervention in AI systems. You will design systems to monitor models, detect unwanted behaviors, prevent misuse, and reduce human intervention. Work across the stack to support real-time safety defenses at scale, collaborating with researchers and policy experts to uphold safety, transparency, and responsible AI.
Location: London
Workplace: Onsite
Employment Type: Full time
Job Function: Software Engineering

Key Responsibilities

  • •Develop the foundational systems which power Safeguards, including infrastructure for data storage and management, metric and evaluation systems, and tooling for human and agentic review.
  • •Ensure the day-to-day running of Safeguards systems and hold a high operational bar which serves both safety and customers while reducing the amount of human intervention and oversight required.
  • •Build robust and reliable multi-layered defenses for real-time improvement of safety mechanisms that work at scale.
  • •Collaborate with researchers and policy experts to align technical implementations with safety principles.
  • •Contribute to monitoring, observability, and tooling to prevent disallowed use of models.

Pay and Benefits

Salary: GBP 255,000 - 325,000 annually

Key Requirements

  • •Bachelor’s degree in Computer Science, Software Engineering or comparable experience.
  • •4-10+ years of experience in a software engineering position
  • •Proficiency in Python
  • •Ability to work across the stack
  • •Strong communication skills and ability to explain complex technical concepts to non-technical stakeholders
Experience:4-10 yearsSafeguardsTrust & safetyAi safety
Education:Bachelor's
Skills:CommunicationTeamworkProblem-solving
Languages:English
Tech Stack:PythonTypeScriptRustJavaScriptClaude Code

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn