Researcher, Safety Oversight

OpenAI
San Francisco
Workplace: OnsiteFull timeUSD 310,000 - 460,000 annuallyFunction: Research & Scientific (R&D)Education: phdSkills: ["Collaboration","Communication","Problem-solving"]

Senior researcher focused on AI safety and oversight for frontier AI models. Sets research directions to maintain safe AGI, develops monitoring models to detect misuse and misalignment, designs red-teaming pipelines, studies human-values reasoning, and collaborates with legal, policy, and cross-functional teams to ensure high safety standards in deployed systems.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
1 year ago

Researcher, Safety Oversight

âś“ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 13 minutes agoStatus: Live

Job Summary

Senior researcher focused on AI safety and oversight for frontier AI models. Sets research directions to maintain safe AGI, develops monitoring models to detect misuse and misalignment, designs red-teaming pipelines, studies human-values reasoning, and collaborates with legal, policy, and cross-functional teams to ensure high safety standards in deployed systems.
Location: San Francisco
Workplace: Onsite
Employment Type: Full time
Job Function: Research & Scientific (R&D)
Seniority: Sr. Manager level

Key Responsibilities

  • •Develop and refine AI monitor models to detect and mitigate known and emerging patterns of misuse and misalignment.
  • •Set research directions and strategies to make our AI systems safer, more aligned, and more robust.
  • •Evaluate and design effective red-teaming pipelines to examine the end-to-end robustness of our safety systems, and identify areas for future improvement.
  • •Conduct research to improve models’ ability to reason about questions of human values, and apply these improved models to practical safety challenges.
  • •Coordinate and collaborate with cross-functional teams, including T&S, legal, policy and other research teams, to ensure that our products meet the highest safety standards.

Pay and Benefits

Salary: USD 310,000 - 460,000 annually
Equity and Bonus:Equity
Perks:Equity

Key Requirements

  • •4+ years of experience in AI safety, especially RLHF, human-AI collaboration, fairness & biases
  • •Ph.D. or other degree in computer science, machine learning, or a related field
  • •4+ years of research engineering experience and proficiency in Python or similar languages
  • •Proven ability to develop and refine AI monitor models to detect and mitigate misuse and misalignment
  • •Experience in safety research and working with large-scale AI systems
Experience:AI safetyMachine learningRLHF
Education:PhD / Doctorate
Skills:CollaborationCommunicationProblem-solving
Tech Stack:PythonRLHFRed-teamingMachine learningML

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor