Model Policy

OpenAI
San Francisco
Workplace: HybridFull timeFunction: Government Affairs & Public PolicySkills: ["Communication","Collaboration","Problem-solving","Written communication","Critical thinking"]

Define and maintain AI safety policies for frontier models, translating risk and harm models into actionable behavioral specifications, evaluation criteria, and safeguards. Collaborate across research, engineering, product, and operations to ensure policies are measurable, scalable, and responsive to real-world risk.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
OpenAI
OpenAI
4 months ago

Model Policy

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 19 minutes agoStatus: Live

Job Summary

Define and maintain AI safety policies for frontier models, translating risk and harm models into actionable behavioral specifications, evaluation criteria, and safeguards. Collaborate across research, engineering, product, and operations to ensure policies are measurable, scalable, and responsive to real-world risk.
Location: San Francisco
Workplace: Hybrid
Employment Type: Full time
Job Function: Government Affairs & Public Policy

Key Responsibilities

  • •Design and maintain model policies across safety-relevant domains, including dual-use, agentic, and emerging frontier-risk areas.
  • •Translate risk and harm models into clear behavioral specifications, evaluation criteria, grading guidance, and system-level safeguards.
  • •Define practical boundaries between beneficial uses of AI and assistance that could materially enable harm, exploitation, misuse, or unsafe outcomes.
  • •Build policy artifacts that support model training, evaluation, and deployment, partnering with safety researchers, engineers, product teams, and other stakeholders to operationalize policy into scalable model behavior and measurable safeguards.
  • •Use red-teaming results, deployment data, model failures, over-refusals, under-refusals, and ambiguous edge cases to improve policy and evaluation quality over time.

Pay and Benefits

Equity and Bonus:Equity
Perks:EquityRelocation

Key Requirements

  • •Experience designing and maintaining model policies in AI safety contexts.
  • •Ability to translate risk and harm models into clear behavioral specifications, evaluation criteria, grading guidance, and safeguards.
  • •Experience collaborating with research, engineering, product, policy, and operations teams.
  • •Strong written and verbal communication to convey complex tradeoffs and policy decisions.
  • •Ability to move across unfamiliar topics, reason from first principles, and turn ambiguity into practical policy.
Experience:AI safetySafety systemsPolicy
Skills:CommunicationCollaborationProblem-solvingWritten communicationCritical thinking

Company Brief

OpenAI
Develops and deploys advanced generative AI models (including ChatGPT and DALL·E) and AI infrastructure, providing APIs and consumer products to accelerate safe AGI for broad benefit.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Scaleup
Valuation: Hectocorn (USD 100B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2015
Glassdoor
Glassdoor: 4.4
WebsiteLinkedInGlassdoor