Senior System Software Engineer - AI Performance and Efficiency Tools

NVIDIA
Santa Clara
Workplace: HybridFull timeUSD 184,000 - 356,500Function: Software EngineeringExperience: 6+ yearsEducation: bachelorsSkills: ["C++","Python","PyTorch","TensorFlow","CUDA","NCCL","Slurm","Kubernetes","GPU","Linux"]

Develop profiling, debugging, and benchmarking tools to optimize AI workloads on NVIDIA GPU clusters. Collaborate with architecture and software teams to translate real-world use cases into features, improve performance and power efficiency, and support AI researchers and HW/SW engineers in scaling workloads.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
3 months ago

Senior System Software Engineer - AI Performance and Efficiency Tools

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Develop profiling, debugging, and benchmarking tools to optimize AI workloads on NVIDIA GPU clusters. Collaborate with architecture and software teams to translate real-world use cases into features, improve performance and power efficiency, and support AI researchers and HW/SW engineers in scaling workloads.
Location: Santa Clara
Workplace: Hybrid
Employment Type: Full time
Job Function: Software Engineering
Seniority: Sr. Manager level

Key Responsibilities

  • •Build internal profiling and analysis tools for AI workloads at large scale.
  • •Build debugging tools for common problems like memory or networking.
  • •Create benchmarking and simulation technologies for AI systems or GPU clusters.
  • •Partner with hardware architects to propose new features or improve existing features with real-world use cases.
  • •Collaborate with multiple global teams to deliver scalable tooling and insights for performance optimization.

Pay and Benefits

Salary: USD 184,000 - 356,500
Equity and Bonus:Equity
Perks:Equity

Key Requirements

  • •BS+ in Computer Science or related field (or equivalent experience) with 6+ years of software development
  • •Strong programming skills in C++ and Python
  • •Good understanding of Deep Learning frameworks like PyTorch and TensorFlow, distributed training and inference
  • •Knowledge of GPU cluster job scheduling (Slurm or Kubernetes), storage and networking
  • •Experience with NVIDIA GPUs, CUDA programming and NCCL
Experience:6+ yearsAIGPUHPCSoftware tooling
Education:Bachelor's
Skills:C++PythonPyTorchTensorFlowCUDANCCLSlurmKubernetesGPULinux
Tech Stack:C++PythonPyTorchTensorFlowCUDANCCLSlurmKubernetesGPULinux

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor