Researcher, Alignment Interpretability
San Francisco
Workplace: OnsiteFull timeUSD 295,000 - 500,000 annuallyFunction: Research & Scientific (R&D)Experience: 2+ yearsEducation: phdSkills: ["Collaboration","Quantitative reasoning","Curiosity","Long-term thinking","Research process focus"]Join the Interpretability team to develop and publish research on techniques for understanding deep neural network representations. You’ll also engineer infrastructure to study model internals at scale, collaborating across teams to pursue projects uniquely suited to OpenAI. The work directly supports OpenAI’s mission by helping ensure future models remain safe as they grow in capability, leveraging mechanistic interpretability and quantitative research rigor.
Loading
Loading job details...
Preparing the role view and application actions.

