Senior Solutions Architect, Cluster Design and Architecture - Networking

NVIDIA
Santa Clara
Workplace: OnsiteFull timeFunction: Solutions Engineering & Sales EngineeringExperience: 8+ yearsEducation: bachelorsSkills: ["Communication","Collaboration","Problem-solving","Presentation","Customer-facing"]

Senior Solutions Architect focused on cluster design and networking for AI/HPC infrastructures. Leads end-to-end network architecture, topology optimization, and performance validation for large GPU clusters, translating complex engineering concepts into customer-ready documentation and guiding field teams and customers through deployments of NVIDIA networking technologies.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
6 months ago

Senior Solutions Architect, Cluster Design and Architecture - Networking

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Senior Solutions Architect focused on cluster design and networking for AI/HPC infrastructures. Leads end-to-end network architecture, topology optimization, and performance validation for large GPU clusters, translating complex engineering concepts into customer-ready documentation and guiding field teams and customers through deployments of NVIDIA networking technologies.
Location: Santa Clara
Workplace: Onsite
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Sr. Manager level

Key Responsibilities

  • •Partner with internal engineering efforts in GPU cluster building and networking and convey architecture and guidelines information both direct to customer and with field teams supporting customers
  • •Guide field teams and their customers in cluster design, weighing design principles but also complex, situational limitations to make the most performant and supportable GPU clusters possible
  • •Work closely with field teams supporting customers to ensure successful first deployments with new products, including new network architectures and topologies
  • •Feedback customer/field perspectives on networking development and workflows back to engineering teams building internal clusters and/or composing customer facing documentation on guidelines and service flows
  • •Perform hands-on work to assist field teams debugging issues relating to network build, configuration, and performance, bringing to bear internal engineering expertise and known bugs

Key Requirements

  • •BS, MS, or PhD in Computer Science, Electrical Engineering, Computer Engineering, Physics, or related field (or equivalent experience)
  • •8+ years of experience in network architecture, network design, network validation and troubleshooting
  • •Proven expertise in designing large-scale distributed systems, AI clusters, or HPC infrastructure
  • •Ability to translate complex engineering concepts into customer-ready documentation, diagrams, and reference material
  • •Experience leading large-scale AI Factory or HPC cluster bring-ups or builds
Experience:8+ yearsAIHPCNetworkingDistributed systemsGPU clusters
Education:Bachelor's
Skills:CommunicationCollaborationProblem-solvingPresentationCustomer-facing
Tech Stack:InfinibandSpectrum-XNVLinkBlueFieldNCCLMPIDistributed trainingGPU clustering

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor