
Hover or tap a row for full statistics (EUR / month on this chart).
Salary analysis
Compared with the selected benchmark ("All roles in London, United Kingdom"), this listing's salary midpoint is about 94% lower. The offer sits below the benchmark range (€3,847–€15,737). The listed pay band (€459–€729) is tighter than the benchmark, which suggests lower salary variability. This benchmark is based on 10 comparable listings.
| Market | Lower bound (25th percentile) | Median | Upper bound (75th percentile) |
|---|---|---|---|
| All roles in London, United Kingdom | €3,847/per month | €7,124/per month | €15,737/per month |
Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US. As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations. Following our Series E funding, Synthesia aims to build human-interactive models that perceive and respond to users, combining text, audio and video in real-time. As a Staff Research Engineer, you will join the Voice team within the 40+ person R&D department, but your scope extends beyond voice. You will define and drive broader vision across teams, own the design and implementation of critical components, and collaborate with senior members across video and research. You will work on voice-to-voice models that produce text and voice simultaneously, enabling natural conversations with back-channeling. Responsibilities include shaping roadmaps, proposing multi-modal architectures, developing streaming conversational systems, designing emotionally expressive interactions, implementing designs from pretraining to post-training, testing architectures, defining evaluation metrics, tracking research in audio-visual diffusion and multimodal LLMs, curating data, leading post-training initiatives, and shipping models to production. You will thrive if you can bring novel ideas, have strong PyTorch experience, time-series knowledge, and a track record of end-to-end model training. Bonus points for real-time architectures, diffusion and neural codec familiarity, and publications.
Job Details
Responsibilities
- Shape our roadmap to create new model capabilities and unlock new functionality for our customer base, on both short and long time horizons.
- Propose novel multi-modal system architectures (especially text and voice).
- Develop and evaluate streaming and conversational systems for low-latency, interactive voice-video synthesis.
Requirements
- Proven experience training deep learning models end-to-end, from data preparation through evaluation.
- Strong general software engineering skills, enabling contributions to a large, shared research infrastructure.
Skills & Technologies

| Location | Active listings |
|---|---|
| London, United Kingdom | 15 |
| Remote - Europe | 14 |
| Remote - Global | 1 |
| Austin, United States | 1 |
| Role type | Active listings |
|---|---|
| Product Designer | 3 |
| Strategic Customer Success Manager | 2 |
| Frontend Engineer | 1 |
| Research Engineer in Data | 1 |
| Staff Research Engineer | 1 |
| Product Manager | 1 |
| Solutions Architect | 1 |
| Application Security Engineer | 1 |
| Engineering Manager | 1 |
| Solutions Consultant | 1 |
| Applied Research Engineer | 1 |
| Senior Brand Designer | 1 |
| Role level | Active listings |
|---|---|
| Senior | 10 |
| Manager | 2 |
| Mid-Level | 2 |
| Expert | 1 |
Related Opportunities
Discover more opportunities that match your interests and skills