Senior Solution Architect, Generative AI - CSP

NVIDIA
Bengaluru, Mumbai
Full timeFunction: Solutions Engineering & Sales EngineeringExperience: 7+ yearsEducation: mastersSkills: ["Communication","Collaboration","Technical leadership","Workshop facilitation"]

Architect end-to-end generative AI solutions focused on LLM training, deployment, and RAG workflows. Partner with customers to translate language-related business challenges into tailored technical solutions, and collaborate with NVIDIA engineering to improve generative AI software. Lead workshops and design sessions, provide technical leadership on training and best practices, and optimize LLM performance on GPU platforms using NVIDIA hardware and software.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
1 day ago

Senior Solution Architect, Generative AI - CSP

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Architect end-to-end generative AI solutions focused on LLM training, deployment, and RAG workflows. Partner with customers to translate language-related business challenges into tailored technical solutions, and collaborate with NVIDIA engineering to improve generative AI software. Lead workshops and design sessions, provide technical leadership on training and best practices, and optimize LLM performance on GPU platforms using NVIDIA hardware and software.
Location: Bengaluru, Mumbai
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Sr. Manager level

Key Responsibilities

  • •Architect end-to-end generative AI solutions with a focus on LLM training, deployment, and RAG workflows.
  • •Collaborate with customers to understand language-related business challenges and design tailored solutions.
  • •Provide feedback and contribute to the evolution of generative AI software with NVIDIA engineering teams.
  • •Engage directly with customers/partners to understand requirements and challenges.
  • •Lead workshops and design sessions and guide training and optimization of large language models using NVIDIA platforms and hardware/software.

Key Requirements

  • •Master's or Ph.D. in Computer Science, Artificial Intelligence, or equivalent experience.
  • •7+ years of hands-on experience in a technical AI role focused on generative AI and LLM training.
  • •Proven track record deploying and optimizing LLM models for inference in production environments.
  • •Expertise in training and fine-tuning LLMs using frameworks such as Megatron-LM, Megatron-Bridge, AutoModel, and PyTorch.
  • •Strong knowledge of GPU cluster architecture and parallel/distributed processing for accelerated training and inference.
Experience:7+ yearsGenerative AILarge language modelsRAGGPU clustersNVIDIA GPU technologies
Education:Master's in Computer Science, Artificial Intelligence
Skills:CommunicationCollaborationTechnical leadershipWorkshop facilitation
Tech Stack:Large Language Models (LLMs)Generative AIRAG workflowsMegatron-LMMegatron-BridgeAutoModelPyTorchGPU cluster architectureParallel processingDistributed computingAWSAzureGCPDockerKubernetesNVIDIA GPU technologies

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor