Research Engineer, Model Evaluations
Anthropic
San Francisco, New York
Workplace: HybridFull timeUSD 320,000 - 485,000 annuallyFunction: Research & Scientific (R&D)Education: bachelorsSkills: ["Communication","Collaboration","Presentation"]Design and run evaluations of Claude’s capabilities across reasoning, agentic behavior, knowledge, and safety; build scalable, reliable evaluation infrastructure; own dashboards to monitor model health; debug anomalous results during training, and collaborate with researchers to translate insights into measurable, defensible metrics that advance AI safety and performance.

