Research Engineer, Code RL (Reinforcement Learning)
Anthropic
San Francisco, New York
Workplace: OnsiteFull-timeUSD 500,000 - 850,000 annuallyFunction: Education & TrainingEducation: bachelorsSkills: ["Communication","Problem-solving","Collaboration","Attention to detail","Adaptability"]Hybrid of research and software engineering focused on building end-to-end RL code for real-world coding tasks. Design RL environments, implement reward signals, run large-scale training on frontier models, diagnose performance, and optimize pipelines to accelerate iteration while ensuring safety and quality.

