Deep Learning Compiler Engineer

NVIDIA
Santa Clara, Austin, Washington, Oregon, Redmond, Texas, California
Full timeUSD 152,000 - 241,500 annuallyFunction: Data Science & Machine LearningExperience: 3+ yearsSkills: ["Independent work","Project scoping","Debugging","Performance analysis","Interpersonal skills"]

Build and optimize CUDA Tile for NVIDIA GPUs, a tile-based programming model shipping with CUDA 13.1. Design and implement compiler transformations, MLIR-based dialects, and lowering passes, and develop public APIs. Ensure tile-based kernels run efficiently across multiple NVIDIA GPU architectures through compiler and performance optimization, including performance analysis and testing in a product-focused team.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
3 days ago

Deep Learning Compiler Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 14 hours agoStatus: Live

Job Summary

Build and optimize CUDA Tile for NVIDIA GPUs, a tile-based programming model shipping with CUDA 13.1. Design and implement compiler transformations, MLIR-based dialects, and lowering passes, and develop public APIs. Ensure tile-based kernels run efficiently across multiple NVIDIA GPU architectures through compiler and performance optimization, including performance analysis and testing in a product-focused team.
Location: Santa Clara, Austin, Washington, Oregon, Redmond, Texas, California
Employment Type: Full time
Job Function: Data Science & Machine Learning

Key Responsibilities

  • •Work on CUDA Tile, a tile-based programming model for NVIDIA GPUs.
  • •Design and implement compiler transformations.
  • •Develop MLIR-based dialects and lowering passes, including performance optimization for tile-based kernels.
  • •Define public APIs and implement compiler/optimization techniques.
  • •Optimize execution efficiency across multiple generations of NVIDIA GPU architectures through analysis, debugging, and testing.

Pay and Benefits

Salary: USD 152,000 - 241,500 annually
Equity and Bonus:Equity

Key Requirements

  • •Bachelors, Masters, or Ph.D. in Computer Science, Computer Engineering, or a related field (or equivalent experience).
  • •3+ years of relevant work or research experience in compiler optimization, performance analysis, and IR design.
  • •Ability to work independently, define project goals and scope, and lead your own development effort.
  • •Excellent C/C++ programming and software design skills, including debugging, performance analysis, and test design.
  • •Strong interpersonal skills and ability to work in a dynamic product-oriented team.
Experience:3+ yearsDeep learningCompilersGPU performance
Education:
Skills:Independent workProject scopingDebuggingPerformance analysisInterpersonal skills
Tech Stack:CUDA TileCUDAC/C++MLIRLLVMXLATVMOpenCLGPU architectureIR designDeep learning modelsGenerative AILarge language modelsRecommendation systemsSpeech recognitionImage classification

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor