Facilities Operations Lead - Compute

OpenAI
San Francisco
Workplace: OnsiteFull timeUSD 159,000 - 240,000 annuallyFunction: Business OperationsExperience: 10+ yearsSkills: ["Leadership","Communication","Problem-solving","Teamwork"]

Lead the commissioning, deployment, and long-term operation of OpenAI’s next-generation AI data centers. Bridge data center construction and hardware deployment, defining commissioning plans and onsite bring-up, owning operations and maintenance for large-scale facilities. Collaborate with design, construction, and hardware teams to codify repeatable processes for new builds, manage monitoring and uptime, staff on-site operations, and define downtime and SLA procedures to ensure reliability.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
4 months ago

Facilities Operations Lead - Compute

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 1 hour agoStatus: Live

Job Summary

Lead the commissioning, deployment, and long-term operation of OpenAI’s next-generation AI data centers. Bridge data center construction and hardware deployment, defining commissioning plans and onsite bring-up, owning operations and maintenance for large-scale facilities. Collaborate with design, construction, and hardware teams to codify repeatable processes for new builds, manage monitoring and uptime, staff on-site operations, and define downtime and SLA procedures to ensure reliability.
Location: San Francisco
Workplace: Onsite
Employment Type: Full time
Job Function: Business Operations

Key Responsibilities

  • •Define and execute sequences of operations, commissioning steps, and bring-up processes for mission-critical data center facilities.
  • •Interface with the design and hardware teams to define deployment procedures tailored to each data center and hardware configuration.
  • •Oversee installation, commissioning, and operational readiness of large-scale data center campuses.
  • •Manage monitoring, maintenance, and quality control of the data center infrastructure, including high-performance liquid cooling systems.
  • •Develop on-site operations staffing strategy and procedures for planned and unplanned downtime and SLAs.

Pay and Benefits

Salary: USD 159,000 - 240,000 annually
Equity and Bonus:Equity
Perks:Equity

Key Requirements

  • •10+ years of experience in large-scale data center facility operations, commissioning, or critical infrastructure engineering.
  • •Deep familiarity with liquid-cooled IT systems, including CDU and in-rack/in-row manifold design, bring-up, and servicing.
  • •Enjoy defining and improving infrastructure processes at the intersection of construction, hardware, and operations.
  • •Comfortable owning infrastructure from deployment to long-term maintenance and failure recovery.
  • •Experience responding to field issues and managing reliability through well-defined operational processes.
Experience:10+ yearsData centerInfrastructure
Skills:LeadershipCommunicationProblem-solvingTeamwork
Tech Stack:Liquid coolingCDUIn-rack manifoldDeployment proceduresMonitoringSLAs

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor