Machine Learning Engineer, API Multicloud

OpenAI
San Francisco
Full timeUSD 295,000 - 445,000 annuallyFunction: Data Science & Machine LearningExperience: 7+ yearsEducation: mastersSkills: ["Python","Rust","PyTorch","TensorFlow","AWS"]

Machine Learning Engineer to build and scale production ML systems for model customization and post-training workflows in OpenAI’s API Multicloud team. You’ll diagnose training and evaluation workflows, collaborate with Research, Applied, Safety Systems, and infrastructure teams, and enable OpenAI models to run in AWS-native environments for enterprise use cases.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
3 months ago

Machine Learning Engineer, API Multicloud

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Machine Learning Engineer to build and scale production ML systems for model customization and post-training workflows in OpenAI’s API Multicloud team. You’ll diagnose training and evaluation workflows, collaborate with Research, Applied, Safety Systems, and infrastructure teams, and enable OpenAI models to run in AWS-native environments for enterprise use cases.
Location: San Francisco
Employment Type: Full time
Job Function: Data Science & Machine Learning

Key Responsibilities

  • •Define target model behaviors with strategic customers and internal teams, diagnose failure modes, and translate real-world needs into training, evaluation, and system requirements.
  • •Build and scale production ML systems for model customization, post-training, and fine-tuning-as-a-service workflows.
  • •Investigate whether training and customization workflows are producing the intended outcomes, and identify changes to data, evaluation, training, or infrastructure that improve performance.
  • •Partner with backend and infrastructure engineers to integrate ML capabilities into AWS-native API environments.
  • •Feed learnings from partner deployments back into the platform by proposing and implementing improvements to post-training systems, tooling, APIs, and developer workflows.

Pay and Benefits

Salary: USD 295,000 - 445,000 annually
Equity and Bonus:Equity

Key Requirements

  • •7+ years of professional engineering experience in relevant ML, infrastructure, or product-driven engineering roles.
  • •Strong ML engineering experience building, training, fine-tuning, evaluating, or deploying production AI systems, with hands-on experience in deep learning, transformer models, and frameworks like PyTorch or TensorFlow.
  • •Familiarity with training and fine-tuning large language models, including methods like supervised fine-tuning, distillation, preference optimization, reinforcement learning, or other post-training techniques.
  • •Strong software engineering fundamentals, including data structures, algorithms, systems design, and high-quality production code in Python, Rust, or similar languages.
  • •Experience with model customization, evaluation systems, data pipelines, distributed systems, cloud infrastructure, or production ML platform tradeoffs.
Experience:7+ yearsAICloudMachine learningOpenAI
Education:Master's
Skills:PythonRustPyTorchTensorFlowAWS
Languages:English
Tech Stack:PythonRustPyTorchTensorFlowAWS

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor