Research Engineer (Reinforcement Learning)
North America, EMEA
Workplace: RemoteFull timeUSD 135,000 - 300,000 annuallyFunction: Education & TrainingSkills: ["Collaboration","Experimenting","Data quality focus","Analytical thinking","Risk-aware design"]Build post-training capabilities for voice and text agents, including training environments, synthetic data pipelines, verifiers, and release evaluations. Own end-to-end training experiments and clearly communicate what improved model behavior. Select and adapt open-weight base models, ensure trained behavior holds up across multi-session voice/text use, and ship models into production while monitoring and iterating from real usage.
Loading
Loading job details...
Preparing the role view and application actions.

