Staff + Senior Software Engineer, Inference Deployment

Anthropic
San Francisco, New York, Seattle
Workplace: OnsiteFull timeUSD 320,000 - 485,000 annuallyFunction: Software EngineeringEducation: bachelorsSkills: ["Kubernetes","Python","Rust","Canary deployments","Blue-green deployments","CLI tools","Web UI","GPU","TPU","Trainium"]

Design and build deployment infrastructure for inference code, orchestrating validation, scheduling, and rollout across GPU/TPU/Trainium fleets. Optimize cycle time from merge to production with capacity-aware scheduling, observability dashboards, and cross-team collaboration to minimize impact on serving capacity and maximize deployment velocity.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
2 months ago

Staff + Senior Software Engineer, Inference Deployment

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 10 hours agoStatus: Live
Reposted: similar role first listed 7 months ago

Job Summary

Design and build deployment infrastructure for inference code, orchestrating validation, scheduling, and rollout across GPU/TPU/Trainium fleets. Optimize cycle time from merge to production with capacity-aware scheduling, observability dashboards, and cross-team collaboration to minimize impact on serving capacity and maximize deployment velocity.
Location: San Francisco, New York, Seattle
Workplace: Onsite
Employment Type: Full time
Job Function: Software Engineering
Seniority: Sr. Manager level

Key Responsibilities

  • •Own deployment orchestration that continuously moves validated inference builds into production across GPU, TPU, and Trainium fleets, unattended under normal conditions
  • •Improve capacity-aware deployment scheduling to maximize deployment throughput against constrained accelerator budgets and variable fleet sizes
  • •Extend deployment observability — dashboards and tooling that answer "what code is running in production," "where is my commit," and "what validation passed for this deploy"
  • •Drive down cycle time from code merge to production with pipeline architectures that minimize serial dependencies and maximize parallelism
  • •Optimize fleet rollout strategies for large-scale deployments across thousands of accelerator chips, minimizing disruption to serving capacity

Pay and Benefits

Salary: USD 320,000 - 485,000 annually

Key Requirements

  • •Strong software engineering skills, including experience designing systems that manage complex state machines and multi-stage pipelines
  • •Proficiency with Kubernetes-based deployments, rolling update mechanics, and container orchestration
  • •Experience building deployment, release, or delivery infrastructure where resource constraints shape the design
  • •A track record of building automation that measurably improves deployment velocity and reliability
  • •Comfort working across the stack — from backend services and databases to CLI tools and web UIs
Experience:AIMachine LearningCloud
Education:Bachelor's
Skills:KubernetesPythonRustCanary deploymentsBlue-green deploymentsCLI toolsWeb UIGPUTPUTrainium
Languages:English
Tech Stack:KubernetesPythonRustDockerCanary deploymentBlue-green deploymentsCLI toolsWeb UIGPUTPUTrainium

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn