Senior GPU Supercomputer Scheduler Engineer

NVIDIA
Santa Clara, Redmond
Workplace: OnsiteFull timeUSD 152,000 - 287,500 annuallyFunction: Solutions Engineering & Sales EngineeringExperience: 5+ yearsEducation: bachelorsSkills: ["Communication","Interpersonal","Customer collaboration"]

Senior GPU Supercomputer Scheduler Engineer to design and implement scheduling for GPU compute clusters supporting demanding deep learning and HPC workloads. You’ll develop batch workload management, optimize AI workloads, and improve the GPU ecosystem, collaborating with ML/AI teams to deliver production-grade, scalable scheduling solutions in a dynamic environment on NVIDIA’s MARS platform.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
6 months ago

Senior GPU Supercomputer Scheduler Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 22 hours agoStatus: Live

Job Summary

Senior GPU Supercomputer Scheduler Engineer to design and implement scheduling for GPU compute clusters supporting demanding deep learning and HPC workloads. You’ll develop batch workload management, optimize AI workloads, and improve the GPU ecosystem, collaborating with ML/AI teams to deliver production-grade, scalable scheduling solutions in a dynamic environment on NVIDIA’s MARS platform.
Location: Santa Clara, Redmond
Workplace: Onsite
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering

Key Responsibilities

  • •Design and develop new scheduling features and add-on services to improve GPU compute clusters across dimensions such as resource usage fairness, GPU occupancy, GPU waste, application resilience, performance and power usage.
  • •Design and develop batch workload management and orchestration services
  • •Provide support to staff and end users to resolve batch scheduler issues
  • •Build and improve the ecosystem around GPU-accelerated computing
  • •Performance analysis and optimizations of deep learning workflows

Pay and Benefits

Salary: USD 152,000 - 287,500 annually

Key Requirements

  • •Bachelor’s degree in Computer Science, Electrical Engineering or related field or equivalent experience
  • •5+ years of work experience
  • •Strong understanding of batch scheduling, preferably with experience in schedulers such as SLURM or K8s batch schedulers (Kueue, Volcano, etc.)
  • •Significant experience in systems programming languages such as C/C++ and Go as well as scripting languages such as Python and bash
  • •Established experience in Linux operating system, environment and tools
Experience:5+ yearsHigh-performance computingAIGPU
Education:Bachelor's
Skills:CommunicationInterpersonalCustomer collaboration
Tech Stack:CC++GoPythonBashLinuxDockerSingularityPodmanSLURMK8sKueueVolcanoPyTorchTensorFlow

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor