Solution Architect - AI Labs

NVIDIA
China
Workplace: OnsiteFull timeFunction: Solutions Engineering & Sales EngineeringExperience: 3+ yearsEducation: phdSkills: ["Communication","Independence"]

Design and deliver production-grade generative AI solutions for enterprise customers using NVIDIA’s software and hardware ecosystem. Partner with customers to accelerate LLM training and inference, implement RAG and agentic inference, and analyze workloads for kernel and model optimization. Provide technical demos and onboarding support across NVIDIA products like CUDA, TensorRT-LLM, and NeMo Framework, while driving industry thought leadership in AI, HPC, and data analytics.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
2 weeks ago

Solution Architect - AI Labs

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 41 minutes agoStatus: Live

Job Summary

Design and deliver production-grade generative AI solutions for enterprise customers using NVIDIA’s software and hardware ecosystem. Partner with customers to accelerate LLM training and inference, implement RAG and agentic inference, and analyze workloads for kernel and model optimization. Provide technical demos and onboarding support across NVIDIA products like CUDA, TensorRT-LLM, and NeMo Framework, while driving industry thought leadership in AI, HPC, and data analytics.
Location: China
Workplace: Onsite
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Mid level

Key Responsibilities

  • •Analyze customers’ latest needs and co-develop accelerated computing solutions with key enterprise accounts.
  • •Support industry accounts by driving research/influencing and new business activities.
  • •Deliver technical projects, demos, and client support tasks as directed by Solution Architecture leadership.
  • •Understand top AI Labs customers’ workloads for LLM training/inference acceleration and application optimization for Agent AI/RAG, including kernel analysis.
  • •Assist customers with onboarding NVIDIA software and hardware solutions, including CUDA, TensorRT-LLM, and NeMo Framework.

Key Requirements

  • •3+ years of experience in research/development/application of machine learning, data analytics, or computer vision workflows.
  • •Strong verbal and written communication skills, with the ability to work independently with minimal daily direction.
  • •Knowledge of AI and large-model industry application trends and hotspots.
  • •Familiarity with large-model training and inference optimization methods and related technology stacks.
  • •C/C++/Python programming experience, plus experience with scale-out cloud and/or HPC architectures for parallel programming.
Experience:3+ years
Education:PhD / Doctorate
Skills:CommunicationIndependence
Tech Stack:Large Language Models (LLMs)RAGAgentic inferenceCUDATensorRT-LLMNeMo FrameworkC/C++PythonDeep LearningHPCParallel programmingKernel analysisKernel optimizationScale-out cloudModel acceleration

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor