Data Center Compute Infrastructure

OpenAI
San Francisco
Workplace: OnsiteFull timeUSD 230,000 - 490,000 annuallyFunction: DevOps, Cloud & InfrastructureSkills: ["Collaboration","Technical judgment","Execution bias","Operational excellence","Problem-solving"]

Help build, scale, and operate OpenAI’s global compute infrastructure that turns frontier AI research into real-world capability. You’ll tackle complex problems across distributed systems, ML infrastructure, GPU clusters, power/cooling, networking, manufacturing, and data center delivery. Partner with cross-functional teams to improve reliability, performance, efficiency, and scalability, and develop tools and processes that accelerate deployment and bring new compute platforms and facilities from concept to production.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
1 month ago

Data Center Compute Infrastructure

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 8 hours agoStatus: Live

Job Summary

Help build, scale, and operate OpenAI’s global compute infrastructure that turns frontier AI research into real-world capability. You’ll tackle complex problems across distributed systems, ML infrastructure, GPU clusters, power/cooling, networking, manufacturing, and data center delivery. Partner with cross-functional teams to improve reliability, performance, efficiency, and scalability, and develop tools and processes that accelerate deployment and bring new compute platforms and facilities from concept to production.
Location: San Francisco
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Help build, scale, and operate OpenAI’s global compute infrastructure.
  • •Solve complex problems across software, hardware, manufacturing supply chain, and data center systems.
  • •Improve reliability, performance, efficiency, and scalability of critical infrastructure.
  • •Partner with cross-functional teams to bring new compute capacity online quickly and reliably.
  • •Identify bottlenecks across technical, operational, and physical systems and develop practical solutions.

Pay and Benefits

Salary: USD 230,000 - 490,000 annually
Equity and Bonus:Equity

Key Requirements

  • •Have experience building, scaling, or operating complex technical systems.
  • •Collaborate across disciplines including software, hardware, operations, and physical infrastructure.
  • •Have strong technical judgment and a bias toward execution.
  • •Care deeply about reliability, speed, safety, and operational excellence.
  • •Work comfortably on ambiguous, high-impact problems where the path forward isn’t always defined.
Experience:AIInfrastructureDistributed systemsMLHigh-performance computing
Skills:CollaborationTechnical judgmentExecution biasOperational excellenceProblem-solving
Tech Stack:Distributed systemsML infrastructureGPU clustersCloud-scale platformsHigh-performance computingPowerCoolingNetworkingData center systems

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor