Manager, AI/HPC Infrastructure Technical Delivery — India

NVIDIA
Pune
Full timeFunction: Transportation & Fleet OperationsExperience: 10+ yearsEducation: mastersSkills: ["Leadership","Coaching","Communication","Customer focus","Problem-solving"]

Lead and manage a technical team of networking (ETH/IB) solutions architects and specialists, owning people management and driving end-to-end delivery of large-scale AI/HPC infrastructure across India and APAC. Provide overall technical control for InfiniBand, HPC clusters, and AI orchestration tool deployments, guiding design, deployment, and operational reliability while partnering with engineering, sales, product, and support to meet customer needs.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
4 days ago

Manager, AI/HPC Infrastructure Technical Delivery — India

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 7 hours agoStatus: Live

Job Summary

Lead and manage a technical team of networking (ETH/IB) solutions architects and specialists, owning people management and driving end-to-end delivery of large-scale AI/HPC infrastructure across India and APAC. Provide overall technical control for InfiniBand, HPC clusters, and AI orchestration tool deployments, guiding design, deployment, and operational reliability while partnering with engineering, sales, product, and support to meet customer needs.
Location: Pune
Employment Type: Full time
Job Function: Transportation & Fleet Operations
Seniority: Manager level

Key Responsibilities

  • •Lead, manage, and develop a technical team (a few to dozens) with full people management lifecycle responsibilities, including hiring, onboarding, performance management, career development, coaching, and retention.
  • •Provide overall technical control for InfiniBand, HPC, and AI cluster delivery projects across India and APAC, ensuring adherence to NVIDIA DC product standards and customer requirements.
  • •Oversee design, deployment, and operational reliability of large-scale AI/HPC infrastructure for new and existing customers, focusing on InfiniBand networking, HPC clusters, and AI orchestration tools.
  • •Drive technical decision-making for complex delivery projects, resolving critical issues across InfiniBand networks, data center architecture, network automation, and GPU-accelerated infrastructure.
  • •Collaborate cross-functionally (engineering, sales, product, support) and act as a senior technical leadership point of contact for customers/partners, communicating delivery status, challenges, and solutions to senior leadership.

Key Requirements

  • •BS/MS/PhD (or equivalent) in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, or a related technical field.
  • •10+ years of progressive data center infrastructure experience, plus 5+ years of people management experience leading a technical team (10+ members).
  • •Strong technical background in NVIDIA DC products (networking solutions, InfiniBand, AI orchestration tools such as BCM) and full-stack data center technologies (storage, server, networking, virtualization, GPU).
  • •Deep expertise in networking fundamentals (TCP/IP stack), data center architecture, and InfiniBand networks with hands-on experience in medium to large-scale HPC/AI environments.
  • •Experience overseeing end-to-end delivery of large-scale AI/HPC infrastructure projects with a focus on performance, reliability, and customer satisfaction.
Experience:10+ yearsData center infrastructureHPCAI/MLNetworkingAPAC
Education:Master's
Skills:LeadershipCoachingCommunicationCustomer focusProblem-solving
Certifications:CCNPCCIEHCIELinux networking CertificationsNVIDIA-related certifications
Tech Stack:EthernetInfiniBandHPCAI orchestrationBase Command Manager (BCM)TCP/IPData center architectureNetwork automationStorageServerNetworkingVirtualizationGPULinuxSlurmPBSMonitoringLoggingAlerting

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor