Model Policy Manager, Agentic Safety
San Francisco
Workplace: HybridFull timeUSD 207,000 - 335,000 annuallyFunction: Government Affairs & Public PolicySkills: ["Communication","Collaboration","Analytical thinking"]Shape how OpenAI identifies and mitigates real-world risks from model misalignment as models become more autonomous. Investigate harmful behavior across long trajectories and translate findings into behavioral policies, evaluations, monitoring, and safeguards. Partner with research, engineering, security, and product teams to balance safety, utility, and business risk, and inform deployment decisions, system cards, and safeguards reports.
Loading
Loading job details...
Preparing the role view and application actions.

