AI Research Intern, TAO Multi-Modal Model Development - 2026

NVIDIA
Hanoi, Ho Chi Minh City
Workplace: OnsiteInternshipFunction: Research & Scientific (R&D)Skills: ["Problem-solving","Attention to detail","Collaborative mindset"]

Extend an AI Research internship for the TAO (Train, Adapt, Optimize) Multi-Modal Model Development project. You’ll develop and fine-tune multi-modal models, contribute to vision-language modeling and universal segmentation systems, and run experiments and benchmarking for accuracy, robustness, and scalability. Working with engineers and researchers in Hanoi/HCM City, you’ll collaborate on research-to-production integration, including pipelines and NVIDIA SDKs, plus code reviews and technical documentation.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
21 hours ago

AI Research Intern, TAO Multi-Modal Model Development - 2026

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 3 hours agoStatus: Live
Reposted: similar role first listed 7 months ago

Job Summary

Extend an AI Research internship for the TAO (Train, Adapt, Optimize) Multi-Modal Model Development project. You’ll develop and fine-tune multi-modal models, contribute to vision-language modeling and universal segmentation systems, and run experiments and benchmarking for accuracy, robustness, and scalability. Working with engineers and researchers in Hanoi/HCM City, you’ll collaborate on research-to-production integration, including pipelines and NVIDIA SDKs, plus code reviews and technical documentation.
Location: Hanoi, Ho Chi Minh City
Workplace: Onsite
Employment Type: Internship
Job Function: Research & Scientific (R&D)
Seniority: Intern level

Key Responsibilities

  • •Develop and fine-tune multi-modal AI models using NVIDIA’s TAO Toolkit and deep learning frameworks.
  • •Contribute to the design and implementation of vision-language models (VLMs) and universal segmentation systems.
  • •Conduct experiments and benchmarking to evaluate model accuracy, robustness, and scalability.
  • •Collaborate with cross-functional teams to integrate research into production-level pipelines and NVIDIA SDKs.
  • •Participate in research discussions, code reviews, and technical documentation to share insights and improve methodologies.

Key Requirements

  • •Currently pursuing a degree in Computer Science, Computer Engineering, or a related field.
  • •Proven experience with machine learning, deep learning, or computer vision model development.
  • •Strong Python programming skills with proficiency in PyTorch or similar frameworks.
  • •Solid understanding of neural network architectures, transformers, and multi-modal learning techniques.
  • •Familiarity with vision-language models, image segmentation, or large-scale pretraining is a strong plus.
Experience:Multi-modal AIComputer visionMachine learningDeep learningLarge-scale pretraining
Education:
Skills:Problem-solvingAttention to detailCollaborative mindset
Tech Stack:TAO ToolkitPythonPyTorchTransformersDeep learning frameworksNVIDIA SDKs

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor