Evaluations - Member of Technical Staff
Simile
San Francisco
Workplace: OnsiteFull timeUSD 200,000 - 400,000 annuallyFunction: Data Science & Machine LearningSkills: ["Analytical reasoning","Hands-on ownership","Rapid prototyping","Data-driven decision-making","Rigor"]Build the measurement layer for Simile’s AI simulations of human behavior, designing evals, metrics, rubrics, datasets, dashboards, and workflows that make simulation quality accurate, trustworthy, and decision-relevant. Partner with modeling teams to diagnose regressions and maintain eval suites, create product/app applied evals, make uncertainty and ground truth legible, and automate end-to-end evaluation systems using agentic coding tools.

