Senior Developer Technology Engineer - Windows AI Platform

NVIDIA
Santa Clara
Workplace: OnsiteFull timeFunction: Software EngineeringExperience: 8+ yearsEducation: bachelorsSkills: ["C/C++","Python","Windows OS","Profiling","Debugging","OSS","GenAI","LLM","CUDA","Nsight","ONNX Runtime","Llama.cpp","GGML","Ollama","Vulkan","DX12"]

Developer Technology Engineer at NVIDIA focusing on local GPU deployment, profiling, and optimization for AI workloads on the RTX platform. You’ll collaborate with internal and partner teams, train developers, enhance OSS tools on Windows, and influence next-gen GPU features while mentoring junior engineers in a fast-paced, research-driven environment.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
6 months ago

Senior Developer Technology Engineer - Windows AI Platform

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Developer Technology Engineer at NVIDIA focusing on local GPU deployment, profiling, and optimization for AI workloads on the RTX platform. You’ll collaborate with internal and partner teams, train developers, enhance OSS tools on Windows, and influence next-gen GPU features while mentoring junior engineers in a fast-paced, research-driven environment.
Location: Santa Clara
Workplace: Onsite
Employment Type: Full time
Job Function: Software Engineering

Key Responsibilities

  • •Work with internal engineering/product teams and external app developers to solve local end-to-end AI GPU deployment challenges on the NVIDIA RTX AI platform.
  • •Apply profiling and debugging tools to analyze GPU-accelerated AI applications and detect suboptimal runtime performance.
  • •Conduct hands-on trainings, develop sample code and host presentations for efficient end-to-end AI deployment on NVIDIA ARM-based SoCs.
  • •Improve Windows LLM & GenAI UX by contributing to OSS projects like GGML, Llama.cpp, Ollama, ONNX Runtime.
  • •Collaborate with GPU driver/architecture teams and NVIDIA research to influence next-generation GPU features based on real-world workflows.
Travel: Low travel

Pay and Benefits

Equity and Bonus:Equity
Perks:Equity

Key Requirements

  • •8+ years of professional experience in local GPU deployment, profiling and optimization.
  • •Bachelor's or Master’s degree or equivalent experience in Computer Science, Engineering, or a related field.
  • •Strong proficiency in C/C++, Python, software design, programming techniques.
  • •Familiarity with and development experience on the Windows operating system.
  • •Experience working with open-source LLM and GenAI software.
Experience:8+ yearsAIGPUEnterprise AI
Education:Bachelor's
Skills:C/C++PythonWindows OSProfilingDebuggingOSSGenAILLMCUDANsightONNX RuntimeLlama.cppGGMLOllamaVulkanDX12
Languages:English
Tech Stack:C/C++PythonWindowsCUDANsightONNX RuntimeLlama.cppGGMLOllamaVulkanDX12

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor