Senior Technical Program Manager, AI Infrastructure and Capacity Operations

NVIDIA
Santa Clara, United States
Workplace: RemoteFull timeUSD 168,000 - 322,000 annuallyFunction: Program & Project Management (PMO)Experience: 7+ yearsEducation: bachelorsSkills: ["Ownership","Written and verbal communication","Influencing without authority","Program mechanics","Turning ambiguity into decisions"]

Own recurring intake, triage, forecasting, and capacity operations for large-scale AI infrastructure programs. Build mechanisms that keep demand, supply, readiness, risks, decisions, and delivery work visible and sequenced. Lead change control and review preparation, track capacity across its lifecycle, manage infrastructure dependencies, and maintain dashboards and decision/risk documentation. Produce clear leadership reporting and influence cross-functional technical and research stakeholders without authority.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NVIDIA
NVIDIA
1 day ago

Senior Technical Program Manager, AI Infrastructure and Capacity Operations

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 18 hours agoStatus: Live

Job Summary

Own recurring intake, triage, forecasting, and capacity operations for large-scale AI infrastructure programs. Build mechanisms that keep demand, supply, readiness, risks, decisions, and delivery work visible and sequenced. Lead change control and review preparation, track capacity across its lifecycle, manage infrastructure dependencies, and maintain dashboards and decision/risk documentation. Produce clear leadership reporting and influence cross-functional technical and research stakeholders without authority.
Location: Santa Clara, United States
Workplace: Remote
Employment Type: Full time
Job Function: Program & Project Management (PMO)
Seniority: Mid level

Key Responsibilities

  • •Own intake, triage, routing, and request-quality standards for accelerator-capacity requests across engineering and research teams.
  • •Run quarterly and annual demand forecasts, change control, review preparation, and follow-through; identify late or conflicting inputs early.
  • •Prepare capacity-planning and allocation reviews using demand, available supply, commitments, readiness, workload timing, and business priorities.
  • •Track capacity through its lifecycle from request/forecast through delivery and productive use, including assignment and readiness.
  • •Coordinate infrastructure dependencies (access, storage, data movement, networking, readiness checks, migration timing) and maintain dashboards and source-data quality.

Pay and Benefits

Salary: USD 168,000 - 322,000 annually
Equity and Bonus:Equity

Key Requirements

  • •BS/MS/PhD in Electrical Engineering, Computer Science, Computer Engineering, or equivalent experience.
  • •7+ years of technical program management or closely related experience in AI/ML platforms, distributed systems, cloud infrastructure, or compute capacity.
  • •Ownership of recurring operational programs with scarce-resource trade-offs, multiple cadences, executive clarity, and overlapping peak periods.
  • •Strong program mechanics for intake, forecasting, review preparation, dependency management, decision/risk records, action closure, and documentation.
  • •Technical proficiency to understand infrastructure constraints, evaluate metrics and source quality, and turn ambiguous cross-functional work into an operating system with clear decisions and issue paths.
Experience:7+ yearsAI/MLDistributed systemsCloud infrastructureCompute capacity
Education:Bachelor's in Electrical Engineering, Computer Science, Computer Engineering
Skills:OwnershipWritten and verbal communicationInfluencing without authorityProgram mechanicsTurning ambiguity into decisions
Tech Stack:AI/MLDistributed systemsCloud infrastructureCompute capacityDashboardsGPUCluster operationsBatch schedulingObservabilityAPI-enabled automation

Company Brief

NVIDIA
Designs and manufactures GPUs, AI accelerators, and system-on-chip products for gaming, data centers, professional visualization, and automotive markets, enabling advanced graphics, AI, and high-performance computing solutions worldwide.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1993
Glassdoor
Glassdoor: 4.3
WebsiteLinkedInGlassdoor