Machine Learning Engineer, Multimodal Perception and Authentication

OpenAI
San Francisco
Workplace: HybridFull timeUSD 342,000 - 399,000 annuallyFunction: Data Science & Machine LearningSkills: ["Cross-functional collaboration","Interdisciplinary problem-solving"]

Help build future AI systems that understand the physical world by developing multimodal perception and authentication methods using signals from cameras, microphones, and other sensors. Work with specialized and larger multimodal models, design data/training/evaluation for real-world conditions, and assess robustness and failure modes. Partner closely with hardware, firmware, software, and product teams to integrate and validate capabilities in real-time or resource-constrained systems in San Francisco.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
1 day ago

Machine Learning Engineer, Multimodal Perception and Authentication

āœ“ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 7 hours agoStatus: Live

Job Summary

Help build future AI systems that understand the physical world by developing multimodal perception and authentication methods using signals from cameras, microphones, and other sensors. Work with specialized and larger multimodal models, design data/training/evaluation for real-world conditions, and assess robustness and failure modes. Partner closely with hardware, firmware, software, and product teams to integrate and validate capabilities in real-time or resource-constrained systems in San Francisco.
Location: San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: Data Science & Machine Learning

Key Responsibilities

  • •Research and develop multimodal perception and authentication methods across visual, audio, and other sensing signals.
  • •Explore how specialized perception models and larger multimodal models can work together.
  • •Design data, training, and evaluation approaches to improve performance in real-world conditions.
  • •Study model behavior, robustness, and failure modes across sensing, data, and deployment environments.
  • •Integrate and validate new capabilities in real-time or resource-constrained systems with partner teams.

Pay and Benefits

Salary: USD 342,000 - 399,000 annually
Equity and Bonus:Equity
Perks:Relocation

Key Requirements

  • •Strong background in computer vision, audio or speech machine learning, multimodal learning, or sensing.
  • •Experience developing specialized machine learning models and/or larger multimodal models.
  • •Ability to design experiments, build evaluations, and investigate model behavior.
  • •Experience working with sensing hardware, real-time systems, or other deployment constraints.
  • •Proficiency in Python and PyTorch, comfortable with C++ or systems integration.
Experience:Computer visionMultimodal learningAudio/speech MLSensing hardwareReal-time systems
Skills:Cross-functional collaborationInterdisciplinary problem-solving
Tech Stack:PythonPyTorchC++

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALLĀ·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor