This evergreen history article uses authoritative archives and official records. Exact dates are used when documented; gradual inventions and rollouts are described as periods rather than being assigned a misleading single birthday.
Quick facts
- Synthesia was founded in 2017 by Victor Riparbelli, Steffen Tjerrild, Matthias Niessner and Lourdes Agapito.
- The company developed a platform that turns written scripts into videos presented by AI-generated avatars.
- Its earliest public attention came from research and demonstrations that transferred speech and facial performance across languages.
- Synthesia focused strongly on enterprise training, internal communication, education and product explanation.
- By 2026, the platform combined avatars, multilingual voices, templates, collaboration and interactive video workflows.
Origins and founding
Synthesia grew from academic and commercial research into computer vision, graphics and speech synthesis. Its founders included specialists connected to University College London and the Technical University of Munich. They saw that video communication was effective but expensive to produce, update and translate, especially for large organizations.
The company was established in London in 2017. Rather than trying to replace cinema, Synthesia concentrated on structured business video: a presenter speaks to the camera, explains a process and can be changed when the script changes. This narrow use case gave the technology a practical market while synthetic media was still unfamiliar.
The product takes shape
Early demonstrations showed how neural rendering could make a performer appear to speak another language. The commercial platform simplified that research into a browser workflow. A user selected an avatar and language, typed or uploaded a script, chose a template and generated a video without a camera crew.
Custom avatars allowed approved employees or presenters to create a reusable digital version of themselves. The platform expanded its stock avatar library, voices and language support, while adding screen recording, brand kits, media libraries, subtitles and review tools for teams.
Technology and major features
Synthesia combines text-to-speech, facial animation, lip synchronization and neural rendering. A generated voice supplies timing and phonetic information, while the avatar system produces matching mouth, face and body motion. Templates organize scenes much like presentation slides, making it easy to replace text or media and regenerate only the affected section.
Enterprise features include access controls, collaboration, translation and analytics. Later products explored more expressive and interactive avatars, allowing a viewer to choose paths or engage with generated presenters. The company also developed consent procedures for custom avatars and policies restricting deceptive or harmful uses.
Growth and wider influence
Large companies adopted Synthesia for onboarding, compliance training, product education and internal announcements. A single course could be translated into many languages without recording every speaker again. Updates that once required a new shoot could be made by editing the script.
The platform helped establish “AI video for work” as a separate category from cinematic text-to-video. Its closest competitors included HeyGen and other avatar services, but Synthesia differentiated itself through enterprise governance, structured learning content and a long-running focus on responsible use.
Challenges, criticism and responsibility
Synthetic presenters can mislead viewers if they appear to represent a real person or organization without permission. The same technology can support impersonation, fake endorsements and political manipulation. Even authorized avatars raise employment and creative-labor questions if companies prefer a reusable digital person to human performers.
Synthesia has published ethical principles, verification processes and content rules, yet no policy is perfect. Users must disclose synthetic presentation where context demands it, protect biometric data and ensure translated scripts preserve meaning. Avatars are also not ideal for every message; sensitive communication often requires a real human presence.
Where it stands in 2026
By 2026, Synthesia had become a mature enterprise platform for generated presenter video. It offered multilingual avatars, templates, collaboration and integrations designed for repeated organizational communication. Its work showed that AI video could create value without generating complex cinematic worlds.
The next phase is likely to involve more expressive avatars, real-time interaction, personalized learning and closer integration with workplace knowledge systems. Trust will remain the deciding factor: organizations need confidence that the avatar is authorized, the script is accurate and viewers understand what they are watching.
Timeline
| Year | Location | Event | Why it mattered |
|---|---|---|---|
| 2017 | London, United Kingdom | Synthesia is founded | Creates a company focused on neural video synthesis. |
| 2018–2020 | Europe and global research community | Early language-transfer and avatar demonstrations spread | Shows the potential of AI-generated presenters. |
| 2020–2022 | Global enterprise market | Browser-based avatar creation and templates expand | Makes camera-free business video practical. |
| 2023–2025 | Global | Expressive avatars, collaboration and governance improve | Strengthens the platform for large organizations. |
| 2026 | Enterprise learning and communication | Interactive and personalized workflows mature | Moves AI avatars toward responsive workplace experiences. |
Frequently asked questions
Who founded Synthesia?
Victor Riparbelli, Steffen Tjerrild, Matthias Niessner and Lourdes Agapito founded Synthesia.
When was Synthesia founded?
The company was founded in 2017 in London.
What is Synthesia used for?
It is commonly used for training, onboarding, internal communication, product education and multilingual presenter videos.
Does Synthesia require consent for custom avatars?
The platform requires identity and consent procedures for legitimate custom-avatar creation and restricts deceptive use.
Final perspective
Synthesia’s history is a reminder that generative video is not only about spectacular scenes. A reliable presenter who can deliver an approved message in many languages can be commercially transformative. The technology is most valuable when efficiency is paired with consent, accuracy and human judgment.