Researcher, Agent Safety, Oversight and System Mitigations
San Francisco
Workplace: HybridFull timeUSD 380,000 - 500,000 annuallyFunction: Research & Scientific (R&D)Skills: ["Rigorous reasoning","Threat modeling","Experimentation","Evaluation design","Evidence-based iteration"]Work on oversight and system-level mitigations that help increasingly capable AI agents act safely and autonomously in real environments. Build practical controls for agent actions, including sandboxing and permission boundaries, and partner with a Codex harness team to productionize them. Red-team agentic systems to measure prevention of data exfiltration and unsafe tool use, while improving the safety–productivity tradeoff through rigorous evaluations and experimentation.
Loading
Loading job details...
Preparing the role view and application actions.

