Engineering Manager, LLM Inference & Deployment at Scale

NVIDIA
United States
Workplace: OnsiteFull timeUSD 224,000 - 431,250 annuallyFunction: Data Science & Machine LearningExperience: 8+ yearsEducation: bachelorsSkills: ["Leadership","Mentoring","Communication","Debugging","Performance analysis"]

Lead engineering activities to productize deep learning models, including planning, mentoring, and executing projects for inference DL workloads. Align with internal partners across business units on roadmap priorities for optimized numerical, analytics, and deep learning algorithms. Coordinate work across global teams and grow a world-class engineering organization. Bring strong expertise in LLM/VLM deployment and inference optimization, along with deep programming and debugging, performance analysis, and test design skills.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
2 months ago

Engineering Manager, LLM Inference & Deployment at Scale

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 45 minutes agoStatus: Live

Job Summary

Lead engineering activities to productize deep learning models, including planning, mentoring, and executing projects for inference DL workloads. Align with internal partners across business units on roadmap priorities for optimized numerical, analytics, and deep learning algorithms. Coordinate work across global teams and grow a world-class engineering organization. Bring strong expertise in LLM/VLM deployment and inference optimization, along with deep programming and debugging, performance analysis, and test design skills.
Location: United States
Workplace: Onsite
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Manager level

Key Responsibilities

  • •Plan, schedule, mentor, and lead team execution of projects and activities, including creating, optimizing, and deploying inference DL workloads.
  • •Collaborate with internal customers to align priorities across business units.
  • •Coordinate projects across different geographic locations.
  • •Own roadmap development interactions for optimized numerical, analytics, and deep learning algorithms and associated R&D duties.
  • •Grow and develop a world-class team; travel to conferences, other sites, or visit customers occasionally.
Travel: Low travel

Pay and Benefits

Salary: USD 224,000 - 431,250 annually
Equity and Bonus:Equity

Key Requirements

  • •BSc (or equivalent experience).
  • •8+ years of related experience, including 3 years of management/leadership experience.
  • •Experience leading multiple software engineering projects.
  • •Strong experience with Large Language Models (LLMs) and Large Visual-Language Models (VLMs).
  • •Excellent programming, debugging, performance analysis, and test design skills.
Experience:8+ yearsDeep learningAILLMVLM
Education:Bachelor's
Skills:LeadershipMentoringCommunicationDebuggingPerformance analysis
Tech Stack:Deep learningGPUsLLMsVLMsTensorRT-LLMVLLMSGLangInferenceNumerical algorithmsJIRAMicrosoft Project

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor