Senior Product Manager - AI Platform Inference

NVIDIA
Santa Clara
Workplace: OnsiteFull timeUSD 168,000 - 327,750 annuallyFunction: Product ManagementExperience: 6+ yearsEducation: mastersSkills: ["Communication","Interpersonal skills","Product strategy","Roadmapping","Collaboration"]

Build and ship products that help developers create high-performance GenAI inference deployments on NVIDIA GPUs. Own product strategy, roadmaps, and go-to-market plans, and collaborate with internal and external developers to drive model optimization improvements. Partner with leadership to align product direction to company strategy while staying close to the evolving inference landscape. Ideal for a product leader combining technical depth with deep developer focus.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
2 hours ago

Senior Product Manager - AI Platform Inference

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Build and ship products that help developers create high-performance GenAI inference deployments on NVIDIA GPUs. Own product strategy, roadmaps, and go-to-market plans, and collaborate with internal and external developers to drive model optimization improvements. Partner with leadership to align product direction to company strategy while staying close to the evolving inference landscape. Ideal for a product leader combining technical depth with deep developer focus.
Location: Santa Clara
Workplace: Onsite
Employment Type: Full time
Job Function: Product Management
Seniority: Sr. Manager level

Key Responsibilities

  • •Create products that help developers build better inference deployments.
  • •Develop product strategy, roadmaps, and go-to-market plans for inference.
  • •Collaborate with internal and external developers to build roadmaps for model optimization software.
  • •Work with leadership to align and drive company strategy.
  • •Stay engaged with the evolving inference landscape to identify key improvements.

Pay and Benefits

Salary: USD 168,000 - 327,750 annually
Equity and Bonus:Equity

Key Requirements

  • •Experience with inference deployment and optimization software (e.g., vLLM, SGLang, FlashInfer, TensorRT-LLM, Triton, Dynamo, TorchAO).
  • •Strong knowledge of GenAI and/or machine learning concepts, especially performance optimization and software delivery.
  • •BS or MS in Computer Science or Computer Engineering, or equivalent experience.
  • •6+ years of technical product management (or similar) experience at a technology company.
  • •Strong communication and interpersonal skills.
Experience:6+ yearsGenerative AIMachine learningDeep learningDeveloper toolsGPU accelerationOpen source
Education:Master's in Computer Science, Computer Engineering (or similar)
Skills:CommunicationInterpersonal skillsProduct strategyRoadmappingCollaboration
Tech Stack:VLLMSGLangFlashInferTensorRT-LLMTritonDynamoTorchAOTensorRTTorchGPU architectureHW/SW co-designPerformance profiling

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor