Software Engineer, Model Evaluation and Improvement
San Francisco
Workplace: OnsiteFull timeUSD 136,435 - 166,754 annuallyFunction: Software EngineeringExperience: 2+ yearsSkills: ["Collaboration","Curiosity","Comfort with ambiguity"]Build datasets, evaluations, and scalable systems that help improve frontier AI models for scientific tasks. Analyze model failure modes by running experiments across leading models, then design and implement pipelines to curate, transform, and validate structured scientific data. Partner with AI labs and work closely with scientists to translate domain judgment into rigorous evaluation criteria that distinguish strong model behavior, all in a fast-evolving area at the intersection of software engineering and biology.
Loading
Loading job details...
Preparing the role view and application actions.

