BestAI Newsroom research note

This evergreen history article uses authoritative archives and official records. Exact dates are used when documented; gradual inventions and rollouts are described as periods rather than being assigned a misleading single birthday.

Quick facts

  • Synthesia was founded in 2017 by Victor Riparbelli, Steffen Tjerrild, Matthias Niessner and Lourdes Agapito.
  • The company developed a platform that turns written scripts into videos presented by AI-generated avatars.
  • Its earliest public attention came from research and demonstrations that transferred speech and facial performance across languages.
  • Synthesia focused strongly on enterprise training, internal communication, education and product explanation.
  • By 2026, the platform combined avatars, multilingual voices, templates, collaboration and interactive video workflows.

Origins and founding

Synthesia grew from academic and commercial research into computer vision, graphics and speech synthesis. Its founders included specialists connected to University College London and the Technical University of Munich. They saw that video communication was effective but expensive to produce, update and translate, especially for large organizations.

The company was established in London in 2017. Rather than trying to replace cinema, Synthesia concentrated on structured business video: a presenter speaks to the camera, explains a process and can be changed when the script changes. This narrow use case gave the technology a practical market while synthetic media was still unfamiliar.

The product takes shape

Early demonstrations showed how neural rendering could make a performer appear to speak another language. The commercial platform simplified that research into a browser workflow. A user selected an avatar and language, typed or uploaded a script, chose a template and generated a video without a camera crew.

Custom avatars allowed approved employees or presenters to create a reusable digital version of themselves. The platform expanded its stock avatar library, voices and language support, while adding screen recording, brand kits, media libraries, subtitles and review tools for teams.

Technology and major features

Synthesia combines text-to-speech, facial animation, lip synchronization and neural rendering. A generated voice supplies timing and phonetic information, while the avatar system produces matching mouth, face and body motion. Templates organize scenes much like presentation slides, making it easy to replace text or media and regenerate only the affected section.

Enterprise features include access controls, collaboration, translation and analytics. Later products explored more expressive and interactive avatars, allowing a viewer to choose paths or engage with generated presenters. The company also developed consent procedures for custom avatars and policies restricting deceptive or harmful uses.

Growth and wider influence

Large companies adopted Synthesia for onboarding, compliance training, product education and internal announcements. A single course could be translated into many languages without recording every speaker again. Updates that once required a new shoot could be made by editing the script.

The platform helped establish “AI video for work” as a separate category from cinematic text-to-video. Its closest competitors included HeyGen and other avatar services, but Synthesia differentiated itself through enterprise governance, structured learning content and a long-running focus on responsible use.

Challenges, criticism and responsibility

Synthetic presenters can mislead viewers if they appear to represent a real person or organization without permission. The same technology can support impersonation, fake endorsements and political manipulation. Even authorized avatars raise employment and creative-labor questions if companies prefer a reusable digital person to human performers.

Synthesia has published ethical principles, verification processes and content rules, yet no policy is perfect. Users must disclose synthetic presentation where context demands it, protect biometric data and ensure translated scripts preserve meaning. Avatars are also not ideal for every message; sensitive communication often requires a real human presence.

Where it stands in 2026

By 2026, Synthesia had become a mature enterprise platform for generated presenter video. It offered multilingual avatars, templates, collaboration and integrations designed for repeated organizational communication. Its work showed that AI video could create value without generating complex cinematic worlds.

The next phase is likely to involve more expressive avatars, real-time interaction, personalized learning and closer integration with workplace knowledge systems. Trust will remain the deciding factor: organizations need confidence that the avatar is authorized, the script is accurate and viewers understand what they are watching.

Timeline

YearLocationEventWhy it mattered
2017London, United KingdomSynthesia is foundedCreates a company focused on neural video synthesis.
2018–2020Europe and global research communityEarly language-transfer and avatar demonstrations spreadShows the potential of AI-generated presenters.
2020–2022Global enterprise marketBrowser-based avatar creation and templates expandMakes camera-free business video practical.
2023–2025GlobalExpressive avatars, collaboration and governance improveStrengthens the platform for large organizations.
2026Enterprise learning and communicationInteractive and personalized workflows matureMoves AI avatars toward responsive workplace experiences.

Frequently asked questions

Who founded Synthesia?

Victor Riparbelli, Steffen Tjerrild, Matthias Niessner and Lourdes Agapito founded Synthesia.

When was Synthesia founded?

The company was founded in 2017 in London.

What is Synthesia used for?

It is commonly used for training, onboarding, internal communication, product education and multilingual presenter videos.

Does Synthesia require consent for custom avatars?

The platform requires identity and consent procedures for legitimate custom-avatar creation and restricts deceptive use.

Final perspective

Synthesia’s history is a reminder that generative video is not only about spectacular scenes. A reliable presenter who can deliver an approved message in many languages can be commercially transformative. The technology is most valuable when efficiency is paired with consent, accuracy and human judgment.

Sources and references