Senior Manager, Product Reliability

NVIDIA
Santa Clara
Workplace: OnsiteFull timeUSD 256,000 - 385,250 annuallyFunction: Solutions Engineering & Sales EngineeringExperience: 10+ yearsSkills: ["Cross-functional collaboration","Communication","Risk assessment","Data-driven decision making","Leadership"]

Lead and build the product reliability engineering organization in Santa Clara, driving an end-to-end reliability strategy across the product lifecycle from concept through field. Own and coordinate the reliability test plan, methodologies, and qualification processes, and spearhead failure analysis and root-cause investigations to mitigate risk. Partner cross-functionally with engineering, validation, manufacturing, operations, and field teams, and work with ODM partners and third-party labs to scale testing capabilities and provide reliability predictions.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
2 months ago

Senior Manager, Product Reliability

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 22 hours agoStatus: Live

Job Summary

Lead and build the product reliability engineering organization in Santa Clara, driving an end-to-end reliability strategy across the product lifecycle from concept through field. Own and coordinate the reliability test plan, methodologies, and qualification processes, and spearhead failure analysis and root-cause investigations to mitigate risk. Partner cross-functionally with engineering, validation, manufacturing, operations, and field teams, and work with ODM partners and third-party labs to scale testing capabilities and provide reliability predictions.
Location: Santa Clara
Workplace: Onsite
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Manager level

Key Responsibilities

  • •Lead a team of Product Reliability Engineers.
  • •Drive end-to-end reliability strategy across the product lifecycle (concept → qualification → production → field).
  • •Coordinate creation and implementation of the reliability test plan, including test methodologies and qualification processes.
  • •Perform failure analysis and root cause investigations, ensuring timely resolution and risk mitigation.
  • •Partner cross-functionally and with ODM partners/third-party labs to resolve issues, improve robustness, and scale testing capabilities; provide reliability predictions.

Pay and Benefits

Salary: USD 256,000 - 385,250 annually
Equity and Bonus:Equity

Key Requirements

  • •Bachelor’s or Master’s degree in Electrical Engineering, Mechanical Engineering, or a related field (or equivalent experience).
  • •10+ years of experience in reliability, qualification, or hardware engineering positions within datacenter or computer sectors.
  • •5+ years of experience managing personnel.
  • •Strong foundation in reliability engineering principles, including accelerated testing, failure analysis, and statistical modeling.
  • •Experience with system-level hardware such as servers, racks, or complex electronic systems.
Experience:10+ yearsReliability engineeringHardware engineeringDatacenterComputer systems
Education:
Skills:Cross-functional collaborationCommunicationRisk assessmentData-driven decision makingLeadership

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor