Staff Research Engineer - Multimodal Generative Modelling
Synthesia
London
Workplace: RemoteFull timeFunction: Research & Scientific (R&D)Skills: ["Prototype quickly","Iterate efficiently"]Build and ship multimodal generative models for real-time interactive voice-video experiences. Join the Voice team within a 40+ person R&D org to define a research roadmap, propose novel text-and-voice architectures, and develop low-latency streaming conversational systems. Work end-to-end from pretraining and post-training (e.g., DPO, fine-tuning, distillation) through dataset curation, evaluation metrics, integration/testing (neural codecs, diffusion, flow-matching), and production deployment.

