Compiler Engineer - AI Inference

NVIDIA
United Kingdom, Cambridge
Workplace: RemoteFull timeFunction: Data Science & Machine LearningEducation: phdSkills: ["Communication","Interpersonal skills","Collaboration","Debugging","Performance analysis","Test design"]

Build and optimize compilers for AI inference and training, with a focus on kernel generation and computational graph optimizations for next-generation NVIDIA GPUs. Tackle complex compilation challenges and help transition breakthroughs into products. Collaborate across software, hardware, and research teams on hardware/software co-design and contribute to scaling datacenter-scale AI workload deployments. Requires strong C/C++ and Python skills, MLIR experience, and deep understanding of LLM inference.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
3 days ago

Compiler Engineer - AI Inference

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 1 hour agoStatus: Live

Job Summary

Build and optimize compilers for AI inference and training, with a focus on kernel generation and computational graph optimizations for next-generation NVIDIA GPUs. Tackle complex compilation challenges and help transition breakthroughs into products. Collaborate across software, hardware, and research teams on hardware/software co-design and contribute to scaling datacenter-scale AI workload deployments. Requires strong C/C++ and Python skills, MLIR experience, and deep understanding of LLM inference.
Location: United Kingdom, Cambridge
Workplace: Remote
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Mid level

Key Responsibilities

  • •Develop compiler technology focused on kernel generation and computational graph optimizations for next-generation NVIDIA GPUs.
  • •Solve complex compilation problems for AI workloads (inference and training) and transition improvements into enterprise and consumer products.
  • •Partner with experts across software, hardware, and research teams to architect and co-design future silicon.
  • •Contribute to advancement and optimization of datacenter-scale AI workload deployments.
  • •Collaborate across teams to improve AI compiler performance and feasibility for real-world systems.

Key Requirements

  • •BS or MS in Computer Science, Computer Engineering, or a related field (or equivalent experience), with PhD strongly preferred.
  • •3+ years of relevant industry experience specializing in compiler optimizations, synthesis, and placement.
  • •Hands-on experience working with MLIR.
  • •Exceptional C/C++ and Python programming, with strong software design, debugging, performance analysis, and test design skills.
  • •Strong communication and collaboration skills in a fast-paced, product-oriented environment.
Education:PhD / Doctorate
Skills:CommunicationInterpersonal skillsCollaborationDebuggingPerformance analysisTest design
Tech Stack:MLIRCC++PythonLLM inference

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor