Researcher, Recursive Self-Improvement Safety
San Francisco
Workplace: OnsiteFull timeUSD 295,000 - 445,000 annuallyFunction: Research & Scientific (R&D)Skills: ["Strategic thinking","Research taste","Prioritization","Clear communication","Collaboration"]Work with Preparedness to anticipate and mitigate loss-of-control and recursive self-improvement risks as frontier AI capability advances. Help build measurements, automated auditing and monitorability approaches, design experiments and evaluations for safety-relevant misalignment, and turn research into institutional safety pipelines. Coordinate verification mechanisms for future safety agreements and strengthen RSI safety cases by identifying mitigation blindspots—moving quickly from prototypes to production-ready practices.
Loading
Loading job details...
Preparing the role view and application actions.

