Manager, Data Center Operations

X AI
Memphis
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Communication","Initiative","Prioritization","Analytical thinking"]

Oversee data center operations to keep SpaceXAI’s AI infrastructure running reliably, including power, cooling, networking, and hardware deployments targeting 99.999% uptime. Lead and develop a team of data center technicians through training and performance management, streamline hardware lifecycles and incident response, and coordinate with AI specialists and external vendors. Track operational metrics, drive energy-efficient sustainability initiatives, and support scalable expansion and vendor-driven preventative maintenance.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
X AI
X AI
16 hours ago

Manager, Data Center Operations

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 3 hours agoStatus: Live

Job Summary

Oversee data center operations to keep SpaceXAI’s AI infrastructure running reliably, including power, cooling, networking, and hardware deployments targeting 99.999% uptime. Lead and develop a team of data center technicians through training and performance management, streamline hardware lifecycles and incident response, and coordinate with AI specialists and external vendors. Track operational metrics, drive energy-efficient sustainability initiatives, and support scalable expansion and vendor-driven preventative maintenance.
Location: Memphis
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Manager level

Key Responsibilities

  • •Manage site operations across power, cooling, networking, and hardware deployments to maintain 99.999% uptime for AI compute systems.
  • •Lead and develop data center operations technicians through training, performance evaluations, and a collaborative environment.
  • •Own hardware lifecycles, incident resolution, and inventory management while refining procedures for precision and consistency.
  • •Coordinate between technicians, AI specialists, and external vendors to integrate new technology and expand capacity.
  • •Track uptime and efficiency metrics, drive energy-efficient sustainability practices, and run preventative maintenance schedules and ticket workflows in Jira.
Travel: Low travel

Key Requirements

  • •5+ years of experience in data center operations or similar critical environments, with 3+ years managing technical teams.
  • •Proven ability to lead teams effectively in fast-paced, high-responsibility settings.
  • •Solid expertise in server hardware, cabling, and data center technologies from setup through lifecycle management.
  • •Experience supporting compute-heavy environments such as AI, machine learning, or high-performance computing.
  • •Familiarity with Jira and collaborative workflows across teams.
Experience:Data center operationsAIMachine learningHigh-performance computing
Skills:CommunicationInitiativePrioritizationAnalytical thinking
Languages:English
Tech Stack:JiraPythonBash

Company Brief

X AI
Develops advanced artificial intelligence models and research aimed at building safe, general AI and understanding the fundamental nature of the universe. Focuses on large-scale AI systems, research publications, and building foundational AI capabilities.
Industry: AI & Machine Learning
Company Size: Medium (51 to 250 employees)
Growth: Early Stage Startup
Headquarters: San Francisco, United States
Founded: 2023
Website