Staff+ Software Engineer, Infrastructure (Distributed Systems)

Anthropic
San Francisco, New York, Seattle
Workplace: OnsiteFull timeUSD 320,000 - 485,000 annuallyFunction: Software EngineeringEducation: bachelorsSkills: ["Communication","Alignment","Incident response","Reliability focus","Mentoring"]

Own and lead complex, multi-month distributed infrastructure projects from ambiguity to production. Make architectural decisions that other engineers build on, drive technical alignment across research and product teams, and translate compute and infrastructure needs into scalable designs. Improve reliability, scalability, and security as usage grows, set infrastructure strategy and standards, and build operational practices like incident response and postmortems while mentoring engineers.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
10 months ago

Staff+ Software Engineer, Infrastructure (Distributed Systems)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 15 hours agoStatus: Live

Job Summary

Own and lead complex, multi-month distributed infrastructure projects from ambiguity to production. Make architectural decisions that other engineers build on, drive technical alignment across research and product teams, and translate compute and infrastructure needs into scalable designs. Improve reliability, scalability, and security as usage grows, set infrastructure strategy and standards, and build operational practices like incident response and postmortems while mentoring engineers.
Location: San Francisco, New York, Seattle
Workplace: Onsite
Employment Type: Full time
Job Function: Software Engineering
Seniority: Mid level

Key Responsibilities

  • •Independently scope and lead complex, multi-month infrastructure projects from ambiguous starting points through to production systems.
  • •Make architectural decisions that shape the infrastructure foundation for other engineers and teams.
  • •Drive alignment on technical direction across multiple teams in ambiguous problem spaces.
  • •Partner with research and product teams to understand infrastructure and compute needs and translate them into technical designs.
  • •Take ownership of reliability, scalability, and security, and set infrastructure strategy and standards; build operational processes like incident response, postmortems, and on-call rotations while mentoring engineers.

Pay and Benefits

Salary: USD 320,000 - 485,000 annually
Perks:Parental LeaveEquityPaid Leave

Key Requirements

  • •Experience designing, building, and operating large-scale distributed systems or infrastructure in production.
  • •Ability to independently scope and deliver complex, ambiguous, multi-month technical projects.
  • •Experience making architectural decisions that other engineers and teams build on.
  • •Strong software engineering fundamentals and proficiency in at least one programming language (e.g., Python, Rust, Go, or Java).
  • •Experience with modern cloud infrastructure, including Kubernetes and infrastructure-as-code, on AWS and/or GCP.
Experience:Distributed systemsInfrastructureCloud infrastructureKubernetesMachine learning infrastructure
Education:Bachelor's
Skills:CommunicationAlignmentIncident responseReliability focusMentoring
Languages:English
Tech Stack:KubernetesAWSGCPInfrastructure-as-codePythonRustGoJavaLinuxEBPFMachine learning infrastructureGPUsTPUsTrainiumNCCL

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn