Senior Research Scientist | Model Steering
DeepL
London
Workplace: HybridFull timeFunction: Data Science & Machine LearningSkills: ["Hands-on execution","Experimentation","Debugging","Mentoring","Communication"]Lead fine-tuning, post-training, model stearability, and reinforcement learning to advance next-generation LLM-based translation models. Fuse human expert and synthetic data to make translation models follow custom user instructions, rules, and context. Build reward and evaluator models, mitigate reward hacking, and expand toward multimodal context. Own the full model lifecycle from prototyping and experiments to production deployment and continuous improvement, mentoring researchers and engineers.

