Research Scientist, Frontier Risk Evaluations

Scale AI
San Francisco, New York
Workplace: OnsiteFull timeUSD 197,400 - 246,750 annuallyFunction: Data Science & Machine LearningExperience: 3+ yearsSkills: ["Communication","Problem-solving","Collaboration","Writing","Presentation"]

Research Scientist focused on frontier AI risk evaluations, designing evaluation measures, harnesses, and datasets to measure and mitigate risks posed by advanced AI systems. Collaborates with government agencies and labs, publishes methodologies for policymakers, and turns research ideas into working prototypes within cross-functional teams.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Scale AI
Scale AI
4 months ago

Research Scientist, Frontier Risk Evaluations

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 1 hour agoStatus: Live

Job Summary

Research Scientist focused on frontier AI risk evaluations, designing evaluation measures, harnesses, and datasets to measure and mitigate risks posed by advanced AI systems. Collaborates with government agencies and labs, publishes methodologies for policymakers, and turns research ideas into working prototypes within cross-functional teams.
Location: San Francisco, New York
Workplace: Onsite
Employment Type: Full time
Job Function: Data Science & Machine Learning

Key Responsibilities

  • •Design and build evaluation harnesses to test AI models and systems for dangerous capabilities and risks.
  • •Collaborate with government agencies or other labs to scope and design evaluations to measure and mitigate risks posed by advanced AI systems.
  • •Publish evaluation methodologies and write technical reports for policymakers.
  • •Develop datasets and measures for Frontier Risk Evaluations and turn research ideas into working prototypes.
  • •Communicate findings and collaborate with cross-functional teams to advance safe, trustworthy AI deployments.

Pay and Benefits

Salary: USD 197,400 - 246,750 annually
Perks:Health InsuranceRetirement BenefitsLearning StipendPaid Leave

Key Requirements

  • •Must have at least three years of experience addressing sophisticated ML problems in research or product development; able to design and instrument ML pipelines and evaluation harnesses.
  • •Track record of published research in machine learning, particularly in generative AI.
  • •Strong written and verbal communication skills to operate in a cross-functional team.
  • •Experience designing and building evaluation harnesses for testing AI models and systems.
  • •Nice to have experience crafting evaluations/benchmarks for LL M technologies or red-teaming/adversarial testing.
Experience:3+ yearsAi safetyMachine learningResearch
Skills:CommunicationProblem-solvingCollaborationWritingPresentation
Languages:English
Tech Stack:ML pipelinesEvaluation harnessesResearch publications

Company Brief

Scale AI
Provides data labeling, annotation, and infrastructure services to accelerate machine learning and AI development. Supplies high-quality training data, tooling, and APIs for customers in autonomous vehicles, mapping, robotics, and enterprise AI applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2016
WebsiteLinkedIn