Senior Manager, AI Infrastructure Operations
Vultr
United States
Workplace: RemoteFull timeUSD 150,000 - 160,000Function: Data Science & Machine LearningExperience: 6-10 yearsSkills: ["Leadership","Mentoring","Communication","Cross-functional collaboration","Execution"]Lead the engineering team responsible for deploying, operating, and optimizing AI compute clusters at scale. Translate AI Infrastructure roadmaps into execution milestones, drive cluster deployments and hardware bring-up, and ensure reliability through monitoring, automation, and continuous improvements. Oversee bare metal and GPU fleet lifecycle operations, manage incident response, and partner across AI/ML, SRE, Networking, Hardware, and Product to deliver performant, continuously improving AI workloads.

