Research Scientist - Multimodal Agent, Consumer Devices

OpenAI
San Francisco
Workplace: HybridFull timeUSD 380,000 - 445,000 annuallyFunction: Research & Scientific (R&D)Skills: ["Communication","Collaboration","Problem-solving"]

We are seeking a Research Engineer/Scientist to join the Future of Computing Research team to advance RLHF and post-training for personalized, multimodal AI systems. You’ll build learning and evaluation foundations for context-aware, adaptive models, design reward models and evaluation pipelines, and collaborate with safety researchers to ensure trustworthy personalization in real-world scenarios.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
4 months ago

Research Scientist - Multimodal Agent, Consumer Devices

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 6 hours agoStatus: Live

Job Summary

We are seeking a Research Engineer/Scientist to join the Future of Computing Research team to advance RLHF and post-training for personalized, multimodal AI systems. You’ll build learning and evaluation foundations for context-aware, adaptive models, design reward models and evaluation pipelines, and collaborate with safety researchers to ensure trustworthy personalization in real-world scenarios.
Location: San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: Research & Scientific (R&D)

Key Responsibilities

  • •Develop RLHF and post-training methods for multimodal models.
  • •Build reward models and preference-learning pipelines for adaptive, personalized model behavior.
  • •Design datasets, rubrics, and evaluation frameworks that capture user preferences, contextual appropriateness, and long-term value in realistic tasks.
  • •Run experiments on policy improvement using explicit feedback, implicit signals, and model-based grading.
  • •Work on long-horizon evaluation problems, where model quality depends not just on a single response but on whether behavior improves outcomes over time.

Pay and Benefits

Salary: USD 380,000 - 445,000 annually
Equity and Bonus:Equity

Key Requirements

  • •Strong background in machine learning research with experience in RLHF, reward modeling, preference optimization, or post-training for large models.
Experience:Multimodal AIPersonalizationRecommender systemsMemoryHuman-in-the-loop evaluation
Skills:CommunicationCollaborationProblem-solving
Tech Stack:RLHFReinforcement learningMultimodal AIPreference learningReward modeling

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor