Senior Engineering Manager, Kernel and Virt

Digital Ocean
Seattle
Workplace: HybridFull timeFunction: Solutions Engineering & Sales EngineeringSkills: ["Leadership","Communication","Problem-solving"]

Lead and scale the Inference Orchestration team building Kubernetes-based AI infrastructure. You’ll drive strategy, execution, and architectural planning for high-throughput scheduling on massive Kubernetes clusters, optimize GPU utilization, ensure reliability, and collaborate with product and other engineering teams to deliver disaggregated AI inference workloads with strong security and fault-tolerance.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Digital Ocean
Digital Ocean
3 months ago

Senior Engineering Manager, Kernel and Virt

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Lead and scale the Inference Orchestration team building Kubernetes-based AI infrastructure. You’ll drive strategy, execution, and architectural planning for high-throughput scheduling on massive Kubernetes clusters, optimize GPU utilization, ensure reliability, and collaborate with product and other engineering teams to deliver disaggregated AI inference workloads with strong security and fault-tolerance.
Location: Seattle
Workplace: Hybrid
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Sr. Manager level

Key Responsibilities

  • •Recruit, mentor, and coach engineers on the team to foster ownership and technical excellence.
  • •Own the team’s project execution, translating business goals into technical roadmaps and on-time delivery.
  • •Collaborate with Product Management and other engineering teams to align priorities and manage dependencies.
  • •Ensure production health, stability, and on-call rotation for all services owned by the team.
  • •Define the technical roadmap and oversee architecture of high-throughput scheduling for large Kubernetes clusters.

Pay and Benefits

Salary: 251,000

Key Requirements

  • •Proven engineering leadership experience managing and growing high-performing teams in a distributed systems or infrastructure domain.
  • •Kubernetes at scale and AI infrastructure domain knowledge.
  • •Hardware-aware optimization including GPU architectures and interconnects, and hardware topology.
  • •Experience balancing performance against cost using principles like Dominant Resource Fairness (DRF).
  • •Systems engineering and security understanding for shared infrastructure and container runtimes.
Experience:CloudAI infrastructureKubernetesDistributed systems
Skills:LeadershipCommunicationProblem-solving
Languages:English
Tech Stack:KubernetesGPUNVIDIANVLinkPCIeNUMACRIUNVIDIA cuda-checkpointKata ContainersGVisorMicroVMsDRF

Company Brief

Digital Ocean
Provides cloud infrastructure and developer-focused cloud services including scalable droplets, managed databases, Kubernetes, object storage, and networking to simplify deploying and managing applications for developers and small-to-medium businesses.
Industry: Cloud Computing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 250M to 500M
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: New York, United States
Founded: 2011
Glassdoor
Glassdoor: 3.8
WebsiteLinkedIn