Senior Applied Research Engineer - Video

Synthesia
London
Workplace: RemoteFull timeFunction: Research & Scientific (R&D)Skills: ["Independent execution","Scientific communication","Experiment design","Fast iteration","Collaboration"]

Build and optimize production-grade foundation models for human-centric video generation. You’ll lead end-to-end research and engineering projects, scaling latent video diffusion models, improving conditioning for pose/emotion/script/camera control, and advancing distributed training and evaluation frameworks. Partnering closely with engineering, you’ll run controlled experiments and improve training stability and inference efficiency for low-latency, high-resolution, cost-effective deployment.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Synthesia
Synthesia
4 days ago

Senior Applied Research Engineer - Video

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 3 hours agoStatus: Live

Job Summary

Build and optimize production-grade foundation models for human-centric video generation. You’ll lead end-to-end research and engineering projects, scaling latent video diffusion models, improving conditioning for pose/emotion/script/camera control, and advancing distributed training and evaluation frameworks. Partnering closely with engineering, you’ll run controlled experiments and improve training stability and inference efficiency for low-latency, high-resolution, cost-effective deployment.
Location: London
Workplace: Remote
Employment Type: Full time
Job Function: Research & Scientific (R&D)
Seniority: Mid level

Key Responsibilities

  • •Own and execute end-to-end research and engineering projects from hypothesis to production impact.
  • •Develop and scale latent video diffusion models for human-centric video generation.
  • •Design conditioning mechanisms to improve control (pose, emotion, script, camera) while preserving fidelity.
  • •Advance distributed training strategies and improve training stability at multi-node scale.
  • •Design evaluation frameworks and optimize inference for low latency, high resolution, and cost efficiency.

Key Requirements

  • •Strong experience training deep learning models at scale.
  • •Strong Python and PyTorch skills.
  • •Hands-on experience with diffusion models (image domain required; video preferred).
  • •Experience with large scale multi-GPU / multi-node training.
  • •Good understanding of distributed training (DDP, FSDP, DeepSpeed or similar).
Skills:Independent executionScientific communicationExperiment designFast iterationCollaboration
Tech Stack:PythonPyTorchCUDADeepSpeedDDPFSDPSequence parallelismAWSSLURMDockerGitHubCI/CD

Company Brief

Synthesia
Provides an AI video generation platform that creates realistic synthetic presenters and video content from text, enabling enterprises to produce scalable training, marketing, and communications videos without cameras or actors.
Industry: AI & Machine Learning
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series C
Headquarters: London, United Kingdom
Founded: 2017
WebsiteLinkedIn