Research Engineer, Machine Learning (Reinforcement Learning)

Anthropic
San Francisco, New York
Workplace: OnsiteFull timeUSD 500,000 - 850,000 annuallyFunction: Data Science & Machine LearningEducation: bachelorsSkills: ["Communication","Problem-solving","Leadership","Teamwork","Analytical thinking"]

Research Engineer in Reinforcement Learning collaborates with researchers and engineers to advance capabilities and safety of large language models, blending research and engineering to implement novel RL approaches, build scalable infrastructure, and contribute to the research direction. Key work includes agentic models via tool use, testing environments, and performance optimizations across distributed GPU workflows.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Anthropic
Anthropic
1 year ago

Research Engineer, Machine Learning (Reinforcement Learning)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 17 hours agoStatus: Live

Job Summary

Research Engineer in Reinforcement Learning collaborates with researchers and engineers to advance capabilities and safety of large language models, blending research and engineering to implement novel RL approaches, build scalable infrastructure, and contribute to the research direction. Key work includes agentic models via tool use, testing environments, and performance optimizations across distributed GPU workflows.
Location: San Francisco, New York
Workplace: Onsite
Employment Type: Full time
Job Function: Data Science & Machine Learning

Key Responsibilities

  • •Architect and optimize core reinforcement learning infrastructure, from clean training abstractions to distributed experiment management across GPU clusters.
  • •Design, implement, and test novel training environments, evaluations, and methodologies for reinforcement learning agents which push the state of the art for the next generation of models.
  • •Drive performance improvements across our stack through profiling, optimization, and benchmarking. Implement efficient caching solutions and debug distributed systems to accelerate both training and evaluation workflows.
  • •Collaborate across research and engineering teams to develop automated testing frameworks, design clean APIs, and build scalable infrastructure that accelerates AI research.
  • •Contribute to research direction and prototype internal tools for productivity and evaluation.

Pay and Benefits

Salary: USD 500,000 - 850,000 annually
Perks:Paid LeaveParental LeaveEquity

Key Requirements

  • •Proficient in Python and async/concurrent programming with frameworks like Trio
  • •Experience with machine learning frameworks (PyTorch, TensorFlow, JAX)
  • •Industry experience in machine learning research
  • •Ability to balance research exploration with engineering implementation
  • •Strong systems design and communication skills
Experience:Reinforcement learningMachine learning
Education:Bachelor's
Skills:CommunicationProblem-solvingLeadershipTeamworkAnalytical thinking
Languages:English
Tech Stack:PythonTrioPyTorchTensorFlowJAXKubernetesDistributed systemsRustC++

Eligibility

Work Authorization:Sponsorship available.

Company Brief

Anthropic
Develops large-scale AI systems and safety research to create reliable, steerable, and interpretable AI assistants and models for commercial and research applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: San Francisco, United States
Founded: 2021
WebsiteLinkedIn