Machine Learning Infrastructure Engineer, Safeguards Research

Anthropic
San Francisco, New York
Workplace: HybridFull timeUSD 350,000 - 500,000 annuallyFunction: DevOps, Cloud & InfrastructureEducation: bachelorsSkills: ["Collaboration","Communication","Debugging","Problem-solving"]

Build and scale the infrastructure and data pipelines that power Safeguards ML research—enabling researchers to train detection methods, evaluate and score results, and select detections for launch. Own end-to-end training/evaluation workflows, create researcher-friendly libraries and command-line tooling, and embed correctness/sanity checks. Improve throughput, cost, and reliability of large-scale inference and scoring while partnering with researchers and engineers to anticipate evolving needs.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
1 month ago

Machine Learning Infrastructure Engineer, Safeguards Research

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 3 hours agoStatus: Live

Job Summary

Build and scale the infrastructure and data pipelines that power Safeguards ML research—enabling researchers to train detection methods, evaluate and score results, and select detections for launch. Own end-to-end training/evaluation workflows, create researcher-friendly libraries and command-line tooling, and embed correctness/sanity checks. Improve throughput, cost, and reliability of large-scale inference and scoring while partnering with researchers and engineers to anticipate evolving needs.
Location: San Francisco, New York
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Build and scale the infrastructure and data pipelines behind Safeguards machine learning research.
  • •Own training, evaluation, and scoring workflows, reducing time between idea and result.
  • •Design tooling and interfaces (libraries and command line tools) that researchers can use directly.
  • •Build correctness and sanity checking into the stack to keep results trustworthy as models and workloads evolve.
  • •Move highest-value research workflows from experiments to reliable, production-grade jobs and improve throughput, cost, and reliability of inference/scoring workloads.

Pay and Benefits

Salary: USD 350,000 - 500,000 annually
Perks:EquityPaid LeaveParental Leave

Key Requirements

  • •Strong software engineering fundamentals with hands-on coding ability, including proficiency in Python.
  • •Experience building and operating data-intensive or distributed systems in production.
  • •Experience building tooling or infrastructure that other engineers or researchers use as a dependency.
  • •Comfort working across the research-to-deployment pipeline from exploratory experiments to production systems.
  • •Ability to debug performance and correctness problems across an unfamiliar stack, with strong written/verbal communication and collaboration.
Experience:Machine learningDistributed systems
Education:Bachelor's in A field relevant to the role as demonstrated through coursework, training, or professional experience
Skills:CollaborationCommunicationDebuggingProblem-solving
Languages:English
Tech Stack:PythonCommand line toolsLanguage modelingTransformersGPUAccelerator programmingInference optimizationExperiment trackingCaching layers

Eligibility

Work Authorization:Sponsorship available.

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn