Synthesia
Visit websiteStaff Research Engineer - Interactive Avatars
Salary not disclosedRemote
- Engineering
- Full time
- Today
About the role
As a Staff Research Engineer, you will join the R&D department to work on cutting-edge challenges in the Generative AI space, specifically focusing on avatar-centric interactive video diffusion models. You will have the opportunity to work on the applied side of research, turning breakthrough ideas into real product capabilities that impact over 60,000 businesses. This role involves collaborating with a team of researchers and engineers to build AI video agents that can think, act, and react like humans.
Responsibilities
- Adapt diffusion models to incorporate diverse conditioning signals such as audio, motion, and interaction cues.
- Develop methods for streaming infinitely long video sequences at real-time rates.
- Work on the perceptual layer of interactive agents, including understanding user audio and generating appropriate contextual reactions.
- Improve lip-sync accuracy, motion realism, and overall visual quality in video diffusion models.
- Build robust evaluation frameworks and test suites to enable continuous quality tracking.
- Collaborate closely with the data team to define data needs and ensure high-quality datasets.
- Stay up to date with research in world models, interactive human/agent modeling, diffusion models, and related areas.
Required skills
- Machine Learning
- Diffusion models
- Computer vision
- PyTorch
- Python
- Git
Nice to have
- Audio-conditioned video diffusion models
- Video DiT architectures
- Model development pipeline
Benefits
- Competitive compensation
- Stock options
- Bonus
- 25 days of annual leave
- Public holidays
About the Company
Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017 and headquartered in London, the company develops products to enhance visual communication and enterprise skill development.