Senior Solutions Architect, Generative AI

NVIDIA
Gurugram
Full timeFunction: Solutions Engineering & Sales EngineeringExperience: 10+ yearsSkills: ["Communication","Collaboration","Technical leadership","Workshop facilitation"]

Architect end-to-end generative AI solutions focused on LLMs, agentic workflows, and RAG-based systems. Partner with customers to translate language-related business challenges into tailored technical solutions, and support pre-sales with demonstrations and presentations. Collaborate with NVIDIA engineering to provide feedback, while leading workshops and driving training/optimization of LLMs using NVIDIA hardware and software for efficient GPU-based inference and deployment.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
1 day ago

Senior Solutions Architect, Generative AI

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

Architect end-to-end generative AI solutions focused on LLMs, agentic workflows, and RAG-based systems. Partner with customers to translate language-related business challenges into tailored technical solutions, and support pre-sales with demonstrations and presentations. Collaborate with NVIDIA engineering to provide feedback, while leading workshops and driving training/optimization of LLMs using NVIDIA hardware and software for efficient GPU-based inference and deployment.
Location: Gurugram
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Mid level

Key Responsibilities

  • •Architect end-to-end generative AI solutions focused on LLMs, agentic workflows, and RAG.
  • •Collaborate with customers to understand language-related business challenges and design tailored solutions.
  • •Support pre-sales activities with technical presentations and demonstrations of LLM and RAG capabilities.
  • •Lead workshops and design sessions to define and refine generative AI solutions, including training and optimization of LLMs using NVIDIA platforms.
  • •Implement strategies and technical guidance for efficient LLM training and RAG-based workflows, including GPU-focused deployment and performance optimization.

Key Requirements

  • •B.Tech., Master’s, Ph.D. in Computer Science, Artificial Intelligence, or equivalent experience.
  • •10+ years of hands-on experience in a technical role focused on generative AI, specifically training Large Language Models (LLMs).
  • •Proven experience deploying and optimizing LLM models for inference in production environments.
  • •Deep understanding of state-of-the-art language models (e.g., GPT-3, BERT) and their architectures.
  • •Experience training and fine-tuning LLMs using frameworks such as TensorFlow, PyTorch, or Hugging Face Transformers.
Experience:10+ yearsGenerative AILLMProduction deployment
Education:
Skills:CommunicationCollaborationTechnical leadershipWorkshop facilitation
Tech Stack:Large Language Models (LLMs)Agentic AIRAGTensorFlowPyTorchHugging Face TransformersGPT-3BERTGPUDockerKubernetes

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor