Senior Manager, Infrastructure Platform Engineering

Crusoe
San Francisco, Sunnyvale
Workplace: OnsiteFull timeUSD 245,000 - 295,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 10+ yearsSkills: ["Leadership","Mentorship","Team building","Cross-functional collaboration","Communication"]

Lead a hands-on, senior engineering team building core platform systems that turn large-scale compute infrastructure into reliable, secure, and efficiently allocatable capacity. Drive roadmaps for capacity/utilization intelligence, lifecycle management, and platform security, while mentoring engineers and aligning with security and production teams to ensure reliability and trust across cloud and on-prem environments.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Crusoe
Crusoe
2 months ago

Senior Manager, Infrastructure Platform Engineering

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Lead a hands-on, senior engineering team building core platform systems that turn large-scale compute infrastructure into reliable, secure, and efficiently allocatable capacity. Drive roadmaps for capacity/utilization intelligence, lifecycle management, and platform security, while mentoring engineers and aligning with security and production teams to ensure reliability and trust across cloud and on-prem environments.
Location: San Francisco, Sunnyvale
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Sr. Manager level

Key Responsibilities

  • •Lead the team responsible for the platform services that abstract underlying infrastructure into reliable, allocatable capacity, and for the systems that track and reconcile state across a large fleet.
  • •Set the technical roadmap across capacity and utilization intelligence, resource lifecycle and state management, and platform security and trust frameworks.
  • •Drive the design of secure, well-instrumented platform systems — from Kubernetes-based orchestration and automation to lower-level system and hardware integration.
  • •Hire, mentoring, and grow a team of infrastructure software engineers; build a high-performing organization from a strong foundation.
  • •Partner with infrastructure, production engineering, and security teams to align platform capabilities with operational reliability, capacity, and trust requirements.

Pay and Benefits

Salary: USD 245,000 - 295,000 annually
Equity and Bonus:Equity
Perks:Health InsuranceDentalVision401kRsusRemote WorkParental LeaveLife InsuranceMeal AllowanceWellness Stipend

Key Requirements

  • •10+ years of experience in infrastructure or systems software development, with at least 3+ years in an engineering leadership role.
  • •Deep expertise in large-scale infrastructure platforms—building services that pool, allocate, and reconcile compute resources at scale.
  • •Strong background with Kubernetes and cloud platforms (GCP, AWS, or Azure)—orchestration, automation, and operating distributed systems in production.
  • •Experience with distributed state management and control systems—modeling resource and system lifecycle, reconciling desired vs. actual state, and handling failure gracefully across a large fleet.
  • •Experience with efficiency, capacity, or performance engineering—characterizing system behavior, identifying bottlenecks, and driving measurable improvements in utilization or availability.
Experience:10+ yearsInfrastructureCloudAI infrastructureData centerEngineering leadership
Skills:LeadershipMentorshipTeam buildingCross-functional collaborationCommunication
Tech Stack:KubernetesGCPAWSAzurePrometheusOpenTelemetryGrafanaCloud platformsInfrastructure softwareOn-premiseAutomationSecurityTelemetry

Company Brief

Crusoe
Builds vertically integrated, energy-first AI infrastructure and purpose-built AI data centers (Crusoe Cloud), leveraging clean/stranded energy to power large-scale GPU compute for AI training and inference.
Industry: Data Centers
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Decacorn (USD 10B+)
Funding: Series E+
Headquarters: Denver, United States
Founded: 2018
Glassdoor
Glassdoor: 3.7
WebsiteLinkedInGlassdoor