AI Computing Development Intern, TensorRT-LLM - 2027

NVIDIA
Shanghai
InternshipFunction: Data Science & Machine LearningEducation: mastersSkills: ["Interpersonal skills","Proactive","Independent work","Written communication","Oral communication"]

Build and optimize TensorRT-LLM inferencing software for NVIDIA’s AI Computing team. You’ll craft robust inference code that scales across multiple platforms, perform performance analysis and tuning, and track advances in AI research to drive feature updates. Collaborate with software, research, and product teams on machine learning inference direction, provide architectural feedback, and publish key results at scientific conferences.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
6 hours ago

AI Computing Development Intern, TensorRT-LLM - 2027

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Build and optimize TensorRT-LLM inferencing software for NVIDIA’s AI Computing team. You’ll craft robust inference code that scales across multiple platforms, perform performance analysis and tuning, and track advances in AI research to drive feature updates. Collaborate with software, research, and product teams on machine learning inference direction, provide architectural feedback, and publish key results at scientific conferences.
Location: Shanghai
Employment Type: Internship
Job Function: Data Science & Machine Learning
Seniority: Intern level

Key Responsibilities

  • •Craft and develop robust inferencing software that can be scaled across multiple platforms for functionality and performance.
  • •Conduct performance analysis, optimization, and tuning.
  • •Follow academic developments in AI and help feature update TensorRT-LLM.
  • •Provide feedback into architecture and hardware design and development.
  • •Collaborate with software, research, and product teams to guide machine learning inferencing and publish key results in scientific conferences.

Key Requirements

  • •Pursue a Masters or higher degree in Computer Engineering, Computer Science, Applied Mathematics, or a related computing-focused field.
  • •Relevant software development experience.
  • •Excellent C/C++ programming and software design skills, including debugging, performance analysis, and test design.
  • •Strong curiosity about AI and awareness of the latest deep learning developments (e.g., LLMs, generative and recommender models).
  • •Experience with deep learning frameworks such as TensorFlow and PyTorch.
Experience:Deep learningAIGPUsLLMsGenerative AI
Education:Master's in Computer Engineering, Computer Science, Applied Mathematics or related computing-focused degree
Skills:Interpersonal skillsProactiveIndependent workWritten communicationOral communication
Languages:English
Tech Stack:TensorRT-LLMCC++TensorFlowPyTorchDeep learningLLMsGenerative models

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor