Amazon Prime Video’s new AI tech matches lips to dubbed audio
Amazon's Prime Video is rolling out a new AI-powered feature that syncs an actor's lip movements to match human-dubbed audio, aiming to make dubbed content look more natural for viewers. The tool combines AI with visual effects technology and is initially limited to the English dub of the German series Maxton Hall, though Amazon intends to extend it to further titles in future. The move follows similar efforts elsewhere in the industry, with Meta and YouTube having recently introduced AI-driven auto-dubbing and lip-syncing tools for content creators.
The lip-sync feature is already live for seasons one and two of Maxton Hall for subscribers worldwide, and will also be applied to the show's third season when it launches on 9 December. This builds on Prime Video's earlier trial of "AI-aided" dubbing in English and Spanish last year, which covered 12 films and TV shows including El Cid: La Leyenda, Mi Mamá Lora and Long Lost.
- Prime Video launches AI lip-sync tech for dubbed shows, starting with Maxton Hall
- Combines AI and visual effects to match mouths to translated audio
- Maxton Hall season 3 arrives 9 December with the feature included
New here? Start with this
Amazon's Prime Video is the video streaming arm of Amazon, offering films and TV series to subscribers. Dubbing is the process of replacing a show's original spoken dialogue with a translated version recorded by different voice actors, which often results in actors' mouths moving out of sync with the words being heard. Amazon has built an AI tool that adjusts an actor's lip movements on screen to better match the dubbed audio, making foreign-language shows look more natural to watch in other languages.
The technology is being tested on Maxton Hall, a German teen drama, for viewers watching the English-dubbed version. This is part of a wider trend across the entertainment and tech industries, with companies such as Meta and YouTube also developing AI tools that automatically dub and lip-sync video content.
The development matters because dubbing quality has long been a sticking point for streaming services trying to sell foreign-language content to international audiences, and AI tools like this could change how film and TV is localised, translated and consumed worldwide, while also raising questions about the role of human dubbing actors and visual effects artists.
Both sides, in good faith
The strongest fair case each way — we don't pick a winner.
The case for
Supporters see this as a genuine improvement to the viewing experience, arguing that mismatched lip movements have long been dubbing's most immersion-breaking flaw and that using AI purely to align visuals with already human-performed audio is a modest, sensible application of the technology. They point out that no dialogue or performance is being synthetically generated, only the mouth movements are adjusted, so it preserves the original actors' and dubbing artists' work while making foreign-language content more accessible and enjoyable for global audiences. From this view, it is a natural evolution of long-established VFX and localisation practices rather than a radical or deceptive use of AI.
The case against
Sceptics worry that even this limited application of AI to alter an actor's recorded likeness sets a troubling precedent, normalising the manipulation of performers' faces without necessarily requiring their explicit, informed consent for each use. They argue it risks devaluing the craft of dubbing and VFX professionals whose skills may be displaced by automation, and that audiences deserve transparency about when they are watching digitally altered footage rather than an unmodified performance. For these critics, the concern is less about this specific feature and more about the broader trajectory of entertainment companies quietly expanding AI's role in reshaping actors' bodies and performances.