Researcher, Interpretability
San Francisco
Workplace: OnsiteFull timeUSD 310,000 - 460,000 annuallyFunction: Research & Scientific (R&D)Experience: 2+ yearsEducation: phdSkills: ["Collaboration","Curiosity","Problem-solving"]Researcher focused on mechanistic interpretability of deep networks, developing and publishing techniques to understand model representations, and building scalable infrastructure to study model internals. Collaborates across teams to pursue safety-centric AI research, guiding directions toward practical usefulness and long-term scalability. Requires strong engineering, quantitative reasoning, and experience with large-scale AI systems to advance OpenAI’s safety goals.
Loading
Loading job details...
Preparing the role view and application actions.

