Inference Engineer, Robotics

OpenAI
San Francisco
Workplace: HybridFull timeUSD 380,000 - 380,000 annuallyFunction: Administration & Executive AssistanceSkills: ["Communication","Collaboration","Leadership","Problem-solving"]

Lead GPU inference engineering for OpenAI's Sora team, focusing on improving model serving efficiency, inference performance, and scalable infrastructure. You will optimize kernel-level data movement, collaborate with research and product teams, and design serving infrastructure to support multimodal foundation models in a hybrid San Francisco environment. You will mentor engineers, help define technical direction, and drive complex initiatives to completion.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
1 year ago

Inference Engineer, Robotics

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 11 minutes agoStatus: Live

Job Summary

Lead GPU inference engineering for OpenAI's Sora team, focusing on improving model serving efficiency, inference performance, and scalable infrastructure. You will optimize kernel-level data movement, collaborate with research and product teams, and design serving infrastructure to support multimodal foundation models in a hybrid San Francisco environment. You will mentor engineers, help define technical direction, and drive complex initiatives to completion.
Location: San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: Administration & Executive Assistance
Seniority: Manager level

Key Responsibilities

  • •Improve model serving, inference performance, and system efficiency.
  • •Drive optimizations from a kernel and data movement perspective to improve system throughput and reliability.
  • •Partner with research and product teams to ensure models perform effectively at scale.
  • •Design, build, and improve critical serving infrastructure to support Sora’s growth and reliability needs.
  • •Navigate ambiguity, set technical direction, and drive complex initiatives to completion.

Pay and Benefits

Salary: USD 380,000 annually
Equity and Bonus:Equity
Perks:Relocation

Key Requirements

  • •Have deep expertise in model performance optimization, particularly at the inference layer.
  • •Have a strong background in kernel-level systems, data movement, and low-level performance tuning.
  • •Are excited about scaling high-performing AI systems that serve real-world, multimodal workloads.
  • •Can navigate ambiguity, set technical direction, and drive complex initiatives to completion.
Experience:AIMultimodalInferenceModel servingKernel-level systems
Skills:CommunicationCollaborationLeadershipProblem-solving
Languages:English
Tech Stack:GPUKernelCUDALinuxDistributed systems

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor