Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100.
Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US.
As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations.
Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow.
Our mission is to develop and optimize video base models that power realistic, controllable, and emotionally expressive synthetic humans at scale.
This is not pure research.
This is applied research with direct product impact.
You will work on advancing training recipes, scaling distributed systems, improving evaluation frameworks, and optimizing inference to ensure our models are high quality, stable, and efficient enough for real-world deployment.
Your work will directly influence models used by tens of thousands of businesses worldwide.
What you’ll doYou will own and execute end-to-end research and engineering projects, from hypothesis to production impact.
This includes:Developing and scaling latent video diffusion models tailored for human-centric video generationDesigning conditioning mechanisms to improve control (pose, emotion, script, camera) without sacrificing fidelityAdvancing dist.