Senior Research Engineer - Enterprise Products

NVIDIA
Washington, Santa Clara
Workplace: RemoteFull timeUSD 192,000 - 356,500 annuallyFunction: Research & Scientific (R&D)Education: bachelorsSkills: ["Communication","Interpersonal skills","Empirical research mindset","Mentoring","Collaboration"]

Develop and evaluate routing policies to optimize LLM traffic across a mixture of model systems. Build and run agentic benchmarks (e.g., Terminal-Bench), convert results into calibration data and routing profiles, and publish design docs and code to open-source communities. Partner across NVIDIA teams to ensure seamless integration with the accelerated serving stack, applying deep learning and empirical research methods to improve generative AI inference.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
1 month ago

Senior Research Engineer - Enterprise Products

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 5 hours agoStatus: Live

Job Summary

Develop and evaluate routing policies to optimize LLM traffic across a mixture of model systems. Build and run agentic benchmarks (e.g., Terminal-Bench), convert results into calibration data and routing profiles, and publish design docs and code to open-source communities. Partner across NVIDIA teams to ensure seamless integration with the accelerated serving stack, applying deep learning and empirical research methods to improve generative AI inference.
Location: Washington, Santa Clara
Workplace: Remote
Employment Type: Full time
Job Function: Research & Scientific (R&D)
Seniority: Mid level

Key Responsibilities

  • •Design and evaluate routing policies for LLM traffic to optimize use of mixture of model systems.
  • •Build and run agentic benchmarks (e.g., Terminal-Bench) to measure algorithm quality and produce calibration data and routing profiles.
  • •Ship design docs, code, and documentation to an open-source repository and contribute to the community.
  • •Collaborate with engineering teams across NVIDIA to integrate software seamlessly with the accelerated serving stack.
  • •Apply research and iteration cycles to improve inference technologies for generative AI needs.

Pay and Benefits

Salary: USD 192,000 - 356,500 annually
Equity and Bonus:Equity

Key Requirements

  • •Bachelor’s or Master’s degree in Computer Science or equivalent experience.
  • •8+ years of industry experience in deep learning frameworks (PyTorch or TensorFlow).
  • •Experience designing or running LLM evaluations/benchmarks, ideally agentic, with statistically sound conclusions.
  • •Strong understanding of modern machine learning and deep neural networks, including NLP or speech recognition.
  • •Strong computer science fundamentals (algorithms/data structures, computational complexity, parallel/distributed computing, system software).
Experience:Generative AIDeep learningLLM evaluationsMachine learningDistributed systems
Education:Bachelor's
Skills:CommunicationInterpersonal skillsEmpirical research mindsetMentoringCollaboration
Tech Stack:PyTorchTensorFlowCUDALLMTerminal-Bench

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor