Technical Program Manager, Model Deployment & Capacity

OpenAI
San Francisco
Workplace: HybridFull timeUSD 257,000 - 445,000 annuallyFunction: Program & Project Management (PMO)Skills: ["Sound judgment","Crisp execution","Communication","Cross-functional alignment","Decision-making under ambiguity"]

Own the operating system for Chat capacity and mainline model deployment, connecting demand forecasting and capacity allocation to model readiness, rollout planning, and post-deployment learning. Drive cross-functional programs across research, post-training, inference, fleet, infrastructure, and product to establish readiness gates, risk reviews, launch coordination, and scalable tooling. Define decision-ready metrics and communications to improve forecast accuracy, utilization, reliability, latency, quality, and user impact.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
2 days ago

Technical Program Manager, Model Deployment & Capacity

āœ“ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 9 hours agoStatus: Live

Job Summary

Own the operating system for Chat capacity and mainline model deployment, connecting demand forecasting and capacity allocation to model readiness, rollout planning, and post-deployment learning. Drive cross-functional programs across research, post-training, inference, fleet, infrastructure, and product to establish readiness gates, risk reviews, launch coordination, and scalable tooling. Define decision-ready metrics and communications to improve forecast accuracy, utilization, reliability, latency, quality, and user impact.
Location: San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: Program & Project Management (PMO)

Key Responsibilities

  • •Own cross-functional programs for Chat capacity forecasting, allocation, headroom planning, and constrained-capacity operations.
  • •Build intake, prioritization, and decision mechanisms linking product demand and model requirements to serving capacity.
  • •Partner with product, research, inference, fleet, and capacity teams to develop scenarios and drive timely allocation decisions.
  • •Lead model deployment readiness and rollout planning, including allocation, launch sequencing, validation, and operational handoffs.
  • •Define readiness gates, risk reviews, rollback criteria, escalation paths, and drive launch coordination through post-launch learning and scalable tooling.

Pay and Benefits

Salary: USD 257,000 - 445,000 annually
Equity and Bonus:Equity
Perks:EquityRelocation

Key Requirements

  • •Have led complex technical programs in infrastructure, distributed systems, capacity planning, model serving, or large-scale deployment environments.
  • •Reason credibly about demand, supply, headroom, reliability, latency, and quality tradeoffs to produce executable plans.
  • •Build operating mechanisms or tooling that replace fragmented, manual workflows with scalable systems and clear ownership.
  • •Operate effectively in high-ambiguity, constrained environments where priorities change and tradeoffs must be explicit.
  • •Use metrics to guide decisions, identify bottlenecks, and demonstrate measurable improvements in throughput, predictability, or reliability.
Experience:Distributed systemsCapacity planningModel servingLarge-scale deployment
Skills:Sound judgmentCrisp executionCommunicationCross-functional alignmentDecision-making under ambiguity

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALLĀ·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor