Data Scientist, Inference Capacity Optimization

OpenAI
San Francisco
Workplace: HybridFull timeUSD 293,000 - 325,000 annuallyFunction: Data Science & Machine LearningExperience: 5+ yearsEducation: mastersSkills: ["Communication","Data-driven decision-making","Experimentation","Causal analysis","Cross-functional collaboration"]

Partner with Capacity Systems Engineering and cross-functional teams to optimize inference capacity across OpenAI’s global GPU fleet. Build statistical and ML models to improve utilization, latency, throughput, and efficiency, and develop forecasting models for inference demand. Analyze production workloads, design experiments and simulations for scheduling and serving strategies, and deliver dashboards and metrics that guide infrastructure investment and customer-focused trade-offs.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
1 month ago

Data Scientist, Inference Capacity Optimization

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 12 hours agoStatus: Live
Reposted: similar role first listed 1 month ago

Job Summary

Partner with Capacity Systems Engineering and cross-functional teams to optimize inference capacity across OpenAI’s global GPU fleet. Build statistical and ML models to improve utilization, latency, throughput, and efficiency, and develop forecasting models for inference demand. Analyze production workloads, design experiments and simulations for scheduling and serving strategies, and deliver dashboards and metrics that guide infrastructure investment and customer-focused trade-offs.
Location: San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Mid level

Key Responsibilities

  • •Build statistical and machine learning models to profile and improve GPU utilization, latency, throughput, and overall fleet efficiency.
  • •Develop forecasting models for inference demand across products, regions, and model families.
  • •Analyze production workloads to identify latency bottlenecks and capacity constraints, and surface optimization opportunities.
  • •Design experiments and simulations to evaluate scheduling policies, serving strategies, and infrastructure tradeoffs.
  • •Build dashboards and operational metrics, and communicate findings with engineering teams and executive leadership to guide data-driven capacity decisions.

Pay and Benefits

Salary: USD 293,000 - 325,000 annually
Equity and Bonus:Equity

Key Requirements

  • •MS or PhD in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or a related quantitative discipline (or equivalent industry experience).
  • •5+ years of experience working in the infrastructure data science space.
  • •Strong expertise in Python and SQL.
  • •Experience building forecasting, optimization, or predictive models.
  • •Strong understanding of experimentation, statistical inference, and causal analysis.
Experience:5+ yearsInfrastructure data science
Education:Master's in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or related quantitative discipline
Skills:CommunicationData-driven decision-makingExperimentationCausal analysisCross-functional collaboration
Tech Stack:PythonSQL

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor