Senior Research Engineer - Interactive Avatars
Synthesia is hiring a Senior Research Engineer - Interactive Avatars in London, United Kingdom. Level rates it ; you can apply on Level.
AI in this role
Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US.
As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations.
Following our recent Series E funding round, where we raised $200 million, our valuation stands at $4 billion. Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow.
About the role
As a Senior Research Engineer, you will join a team of 40+ Researchers and Engineers within the R&D Department working on cutting edge challenges in the Generative AI space, with a focus on avatar-centric interactive video diffusion models. Within the team you’ll have the opportunity to work on the applied side of our research efforts and directly impact our solutions that are used worldwide by over 60,000 businesses.
This is a unique opportunity for experts in machine learning and diffusion models to shape the future of AI video agents that can think, act, and react like humans. As part of our Interactive Avatars Team, you’ll work on cutting-edge research with a clear focus on turning breakthrough ideas into real product capabilities. You’ll join a team that moves fast, iterates often, and builds models that ship and make a meaningful impact. Example tasks and responsibilities include:
Adapt diffusion models to incorporate diverse conditioning signals (e.g., audio, motion, interaction cues).
Develop methods for streaming infinitely long video sequences at real-time rates.
Work on the perceptual layer of interactive agents, including understanding user audio and generating appropriate contextual reactions.
Improve lip-sync accuracy, motion realism, and overall visual quality in video diffusion models.
Build robust evaluation frameworks and test suites to enable continuous quality tracking.
Collaborate closely with our data team to define data needs and ensure high-quality datasets.
Stay up to date with research in world models, interactive human/agent modeling, diffusion models, and related areas.
What we are looking for:
Comfortable owning and executing on the responsibilities listed above.
Strong ML (e.g., diffusion, GANs, VAEs) and computer vision background with relevant industry experience.
Hands-on experience with diffusion models (ideally avatar-centric or video-focused) and up to date with recent advances.
Proficient in PyTorch and familiar with modern ML frameworks and tooling.
Strong Python engineering skills, confident with git and version control, and a commitment to clean, maintainable research code.
Outcome-driven, detail-oriented, and motivated to push state-of-the-art research into real product impact.
Clear communicator of hypotheses, experiments, and results.
What will make you stand out:
Experience with audio-conditioned video diffusion models and deep knowledge of recent video DiT architectures.
Demonstrated ability to own the full model development pipeline end to end, from data preparation to model design, training, and evaluation.
A strong publication record in areas such as world models, interactive agents, or video diffusion models.
The good stuff...
Attractive compensation
Hybrid work setting with an office in London, Zurich, and Munich
25 days of annual leave + public holidays
Work in a great company culture with the option to join regular planning and socials at our hubs
A generous referral scheme when you know people that are amazing for us
Strong opportunities for your career growth
You can see more about Who we are and How we work here: https://www.synthesia.io/careers
How we rate this
Senior Research Engineer - Interactive Avatars at Synthesia rates 94 out of 100 for how much of the daily work is AI. That makes it Builds AI (AI Level 4 of 4). The level is about AI in the job, not seniority.
Builds AI. The job is building AI systems.
- ●●●● Builds AI80 to 100
- ●●●○ Works on AI60 to 79
- ●●○○ Uses AI40 to 59
- ●○○○ Little AI0 to 39
Levels come from how often the tools, models and workflows of the role are named in the posting itself. Open the description and count.
Prepare for this job
A free preview built only from this posting: what it asks for, what you could be asked in an interview, and how to adjust your resume.
Skills and AI tools this role asks for
Questions you could be asked
- Walk me through a computer vision problem you solved, from raw data to a deployed model.
- Walk me through how you've used PyTorch in your day-to-day work.
- How would you decide a model or AI system is ready to ship?
- Tell me about a time a model underperformed in production. How did you find out, and what did you change?
Adapt your resume
- List these exact terms on your resume: Computer vision and PyTorch. An applicant tracking system matches the wording, not the idea.
- Attach one line of real, concrete experience to at least one of them. A tool named with nothing behind it rarely survives a human read.
- Lead with what you built, trained or shipped. This role is judged on the AI system itself, not the tools around it.
Want an expert to read your CV for this job?
Free. Send your CV and the role you want next. We reply by email within 2 to 4 business days.
Get new remote research jobs (Builds AI ●●●●) by email
One email a week with the new remote research jobs (Builds AI ●●●●), each rated for how much AI is in the work. No recruiter spam, unsubscribe in one click.
Free. One email a week. Unsubscribe in one click.
Similar roles
Research roles that build AI, at other companies.
What kind of AI work fits you?
Answer 12 practical questions in about three minutes. Get a simple profile, the work it points to, and live roles to explore next.
Find my next step