Senior AI Software Engineer, Kernel Libraries

NVIDIA
Santa Clara
Workplace: RemoteFull timeUSD 184,000 - 287,500 annuallyFunction: Software EngineeringExperience: 6+ yearsEducation: bachelorsSkills: ["Communication","Problem-solving"]

Lead development of AI inference software, designing and optimizing GPU kernels and libraries for NVIDIA’s AI systems. You’ll build abstractions for LLM serving, create JIT compilers and runtimes, and collaborate with CUDA architects, DL frameworks, and open-source communities to accelerate high-impact AI workloads.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
4 months ago

Senior AI Software Engineer, Kernel Libraries

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Lead development of AI inference software, designing and optimizing GPU kernels and libraries for NVIDIA’s AI systems. You’ll build abstractions for LLM serving, create JIT compilers and runtimes, and collaborate with CUDA architects, DL frameworks, and open-source communities to accelerate high-impact AI workloads.
Location: Santa Clara
Workplace: Remote
Employment Type: Full time
Job Function: Software Engineering
Seniority: Manager level

Key Responsibilities

  • •Innovating and developing new AI systems technologies for efficient inference
  • •Designing, implementing, and optimizing kernels for high impact AI workloads
  • •Designing and implementing extensible abstractions for LLM serving engines
  • •Building efficient just-in-time domain specific compilers and runtimes
  • •Collaborating closely with other engineers at NVIDIA across deep learning frameworks, libraries, kernels, and GPU arch teams

Pay and Benefits

Salary: USD 184,000 - 287,500 annually
Equity and Bonus:Equity
Perks:Equity

Key Requirements

  • •Masters degree in Computer Science, Electrical Engineering, or related field (or equivalent experience); PhD are preferred
  • •6+ years (academic/ industry) experience with ML/DL systems development preferable
  • •Strong experience in developing or using deep learning frameworks (e.g. PyTorch, JAX, TensorFlow, ONNX, etc) and ideally inference engines and runtimes such as vLLM, SGLang, and MLC
  • •Strong Python and C/C++ programming skills
  • •Background in domain specific compiler and library solutions for LLM inference and training (e.g. FlashInfer, Flash Attention)
Experience:6+ yearsAIMLDLInference enginesGPU kernels
Education:Bachelor's
Skills:CommunicationProblem-solving
Tech Stack:PythonC++PyTorchJAXTensorFlowONNXVLLMSGLangMLCCUDACuTileTritonFlashInferFlash AttentionMLIRTVM

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor