Staff Machine Learning Research Engineer, Agent Post-training - Enterprise GenAI

Scale AI
San Francisco, New York
Workplace: OnsiteFull timeUSD 218,400 - 273,000 annuallyFunction: Data Science & Machine LearningExperience: 5+ yearsEducation: mastersSkills: ["Communication","Problem-solving","Collaboration"]

Senior research-oriented role developing an agent post-training RL platform for enterprise GenAI. You will train state-of-the-art LLMs, integrate cutting-edge post-training methods, and design multi-agent learning with reward-based outcomes for enterprise use-cases, contributing to a team delivering next-gen AI technologies for customers.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Scale AI
Scale AI
10 months ago

Staff Machine Learning Research Engineer, Agent Post-training - Enterprise GenAI

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 1 hour agoStatus: Live

Job Summary

Senior research-oriented role developing an agent post-training RL platform for enterprise GenAI. You will train state-of-the-art LLMs, integrate cutting-edge post-training methods, and design multi-agent learning with reward-based outcomes for enterprise use-cases, contributing to a team delivering next-gen AI technologies for customers.
Location: San Francisco, New York
Workplace: Onsite
Employment Type: Full time
Job Function: Data Science & Machine Learning

Key Responsibilities

  • •Train state of the art models, developed both internally and from the community, to deploy to our enterprise customers.
  • •Research cutting edge algorithms to integrate directly into our training stack.
  • •Design solutions that enable complex multi-agent systems to directly learn from both process + outcome based rewards.

Pay and Benefits

Salary: USD 218,400 - 273,000 annually
Perks:Health InsuranceDentalVisionRetirement BenefitsLearning BudgetPaid LeaveCommuter Benefits

Key Requirements

  • •5+ years of LLM training in a production environment
  • •Experience with post-training methods like RLHF/RLVR and related algorithms like PPO/GRPO
  • •Publications in top conferences such as NEURIPS, ICLR, or ICML within the last two years
  • •PhD or Masters in Computer Science or a related field
Experience:5+ yearsEnterprise AIGenAILLM training
Education:Master's in Computer Science
Skills:CommunicationProblem-solvingCollaboration
Languages:English
Tech Stack:LLMRLHFRLVRPPOGRPO

Company Brief

Scale AI
Provides data labeling, annotation, and infrastructure services to accelerate machine learning and AI development. Supplies high-quality training data, tooling, and APIs for customers in autonomous vehicles, mapping, robotics, and enterprise AI applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2016
WebsiteLinkedIn