Model Policy Manager, Multimodal Safety

OpenAI
San Francisco
Workplace: HybridFull timeUSD 266,000 - 335,000 annuallyFunction: Government Affairs & Public PolicySkills: ["Judgment","Data-driven thinking","Technical fluency","Collaboration","Consensus building"]

Own and evolve safety policies for multimodal frontier models, translating theories of harm and threat models into behavioral rules, evaluation criteria, grading guidance, and safeguards. Detect safety regressions and failure patterns, and develop policy artifacts that support training, evaluation, and deployment. Partner with AI researchers, domain experts, and product teams to operationalize policy into measurable model behavior while working hands-on with model data and evaluation results.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
20 hours ago

Model Policy Manager, Multimodal Safety

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 14 hours agoStatus: Live

Job Summary

Own and evolve safety policies for multimodal frontier models, translating theories of harm and threat models into behavioral rules, evaluation criteria, grading guidance, and safeguards. Detect safety regressions and failure patterns, and develop policy artifacts that support training, evaluation, and deployment. Partner with AI researchers, domain experts, and product teams to operationalize policy into measurable model behavior while working hands-on with model data and evaluation results.
Location: San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: Government Affairs & Public Policy
Seniority: Manager level

Key Responsibilities

  • •Design and maintain model policies for audio, image, video, and omni-modal behavior.
  • •Translate theories of harm and threat models into behavioral safety policies, evaluation criteria, grading guidance, and safeguards.
  • •Identify and analyze safety regressions and failure patterns to find gaps in existing policies and inform policy iteration.
  • •Develop policy artifacts supporting model training, evaluation, and deployment (including behavior instructions, human-data campaigns, golden sets, and evaluations).
  • •Partner with AI researchers, domain experts, and product teams to operationalize policy into measurable model behavior.

Pay and Benefits

Salary: USD 266,000 - 335,000 annually

Key Requirements

  • •Strong judgment about real-world risks of advanced multimodal AI systems.
  • •Experience turning ambiguous safety questions into clear, data-driven policies, behavioral boundaries, and measurable evaluation criteria.
  • •Ability to treat policy as an end-to-end measurable system by testing intended model behavior and diagnosing gaps across policy, data, graders, and safeguards.
  • •Technical fluency to design safety policies for multimodal model behavior that can be trained, measured, and supervised at scale.
  • •Hands-on experience driving consensus and action in ambiguous spaces.
Experience:Multimodal AIFrontier modelsModel safetyAI safety evaluation
Skills:JudgmentData-driven thinkingTechnical fluencyCollaborationConsensus building

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor