Researcher, Agent Safety, Training and Evaluations

OpenAI
San Francisco
Workplace: HybridFull timeUSD 380,000 - 500,000 annuallyFunction: Research & Scientific (R&D)Skills: ["Judgment","Comfort with ambiguity","Independent execution","Collaboration","Technical execution"]

Work on OpenAI’s Agent Safety mission by training and evaluating frontier AI agents to reduce harmful or misaligned actions. Build measurement, data-processing, and evaluation systems that convert real incidents into repeatable safety signals. Collaborate with training, capabilities, oversight, and pre-training partners to ship research-backed mitigations into large-scale agent systems.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
2 days ago

Researcher, Agent Safety, Training and Evaluations

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 9 hours agoStatus: Live

Job Summary

Work on OpenAI’s Agent Safety mission by training and evaluating frontier AI agents to reduce harmful or misaligned actions. Build measurement, data-processing, and evaluation systems that convert real incidents into repeatable safety signals. Collaborate with training, capabilities, oversight, and pre-training partners to ship research-backed mitigations into large-scale agent systems.
Location: San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: Research & Scientific (R&D)

Key Responsibilities

  • •Train and evaluate frontier models to reduce harmful or misaligned agent actions, forming clear hypotheses and executing through ambiguity.
  • •Mine incidents and build scalable measurement, data-processing, and evaluation systems that turn real failures into repeatable safety signals.
  • •Collaborate closely with post-training, capabilities, oversight, and pre-training partners to ship research-backed mitigations into large-scale training and agent systems.

Pay and Benefits

Salary: USD 380,000 - 500,000 annually
Equity and Bonus:Equity
Perks:Relocation

Key Requirements

  • •Demonstrated strength in research engineering, ML engineering, quantitative research, or applied model research, with the ability to own ambiguous projects end to end.
  • •Excellent technical execution across experimentation, data, evaluation, and/or infrastructure.
  • •Strong intuition for modern frontier-model research.
  • •Motivation to work on urgent, practical agent-safety problems even if prior work was outside safety or alignment.
  • •Comfort with ambiguity and excellent judgment for independent execution.
Experience:AI safetyFrontier modelsMachine learningResearch engineeringQuantitative research
Skills:JudgmentComfort with ambiguityIndependent executionCollaborationTechnical execution

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor