Research Engineer, Production Model Post-Training

Anthropic
San Francisco, New York, Seattle
Workplace: OnsiteFull timeUSD 350,000 - 500,000 annuallyFunction: Manufacturing & Production OperationsEducation: bachelorsSkills: ["Communication","Collaboration","Problem-solving","Debugging","Time-management"]

Research Engineer on the Post-Training team, you will train base models through the complete post-training stack to deliver production Claude models. You’ll implement, scale, and improve post-training techniques (e.g., Constitutional AI, RLHF), build robust pipelines for fine-tuning and evaluation, and collaborate with researchers to deploy production-ready methods that enhance quality, safety, and capabilities.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
1 year ago

Research Engineer, Production Model Post-Training

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 9 hours agoStatus: Live

Job Summary

Research Engineer on the Post-Training team, you will train base models through the complete post-training stack to deliver production Claude models. You’ll implement, scale, and improve post-training techniques (e.g., Constitutional AI, RLHF), build robust pipelines for fine-tuning and evaluation, and collaborate with researchers to deploy production-ready methods that enhance quality, safety, and capabilities.
Location: San Francisco, New York, Seattle
Workplace: Onsite
Employment Type: Full time
Job Function: Manufacturing & Production Operations

Key Responsibilities

  • •Implement and optimize post-training techniques at scale on frontier models.
  • •Conduct research to develop and optimize post-training recipes that directly improve production model quality.
  • •Design, build, and run robust, efficient pipelines for model fine-tuning and evaluation.
  • •Develop tools to measure and improve model performance across various dimensions.
  • •Collaborate with research teams to translate emerging techniques into production-ready implementations.

Pay and Benefits

Salary: USD 350,000 - 500,000 annually
Perks:Paid LeaveParental LeaveEquity

Key Requirements

  • •Proficiency in Python, deep learning frameworks, and distributed computing is required.
  • •Strong software engineering skills with experience building complex ML systems.
  • •Experience with training, fine-tuning, or evaluating large language models.
  • •Comfortable working with large-scale distributed systems and high-performance computing.
  • •Ability to balance research exploration with engineering rigor and operational reliability.
Experience:AIMachine learningProduction systems
Education:Bachelor's
Skills:CommunicationCollaborationProblem-solvingDebuggingTime-management
Languages:English
Tech Stack:PythonDistributed computingDeep learningLLMsRLHFConstitutional AI

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn