Site Reliability Engineer (Manufacturing Infrastructure)

SpaceX
United States
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 3+ yearsEducation: bachelorsSkills: ["Ownership","Clear communication","Stakeholder management","Proactive maintenance mindset","Collaboration"]

Own and scale the compute, storage, and networking that keep SpaceX manufacturing systems running for Starship, Starlink, Starshield, and Terafab. Deploy and operate reliable infrastructure with infrastructure-as-code and observability, design for stability and performance, and proactively manage capacity and lifecycle to prevent incidents. Partner with software and manufacturing teams, provide user support, and participate in on-call and site travel as needed.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
SpaceX
SpaceX
2 days ago

Site Reliability Engineer (Manufacturing Infrastructure)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 7 hours agoStatus: Live

Job Summary

Own and scale the compute, storage, and networking that keep SpaceX manufacturing systems running for Starship, Starlink, Starshield, and Terafab. Deploy and operate reliable infrastructure with infrastructure-as-code and observability, design for stability and performance, and proactively manage capacity and lifecycle to prevent incidents. Partner with software and manufacturing teams, provide user support, and participate in on-call and site travel as needed.
Location: United States
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Deploy, upgrade, operate, maintain, and scale compute, storage, and networking for manufacturing systems across Starship, Starlink, Starshield, and Terafab
  • •Manage infrastructure as code and use observability to understand platform health
  • •Design for reliability, stability, and scale; remove bottlenecks using measurement and engineering
  • •Practice proactive maintenance including capacity planning and lifecycle management while reducing toil
  • •Participate in on-call and travel to sites as needed for deployments and incidents
Travel: Medium travel

Key Requirements

  • •Bachelor’s degree in computer science, information systems, or an engineering discipline, or 3+ years of professional experience in SRE or DevOps in lieu of a degree
  • •1+ years of software development experience
  • •Experience with Linux operating systems
  • •Ability to deploy, upgrade, operate, maintain, and scale compute, storage, and networking systems in production
  • •Experience translating high-level requirements into implementations from first principles
Experience:3+ yearsInfrastructureProduction systemsInfrastructure as code
Education:Bachelor's
Skills:OwnershipClear communicationStakeholder managementProactive maintenance mindsetCollaboration
Tech Stack:LinuxTerraformAnsiblePuppetDockerKubernetesVSphereQEMUKVMPostgresClickhouse

Eligibility

Nationality:US National

Company Brief

SpaceX
Designs, manufactures, and launches advanced rockets and spacecraft for commercial and government customers, aiming to reduce space transportation costs and enable human life on Mars through reusable launch vehicles and integrated space systems.
Industry: Aerospace Manufacturing
Company Size: Enterprise (1,001+ employees)
Growth: Established Company
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: Hawthorne, United States
Founded: 2002
Glassdoor
Glassdoor: 4.2
WebsiteLinkedIn