Prime Video Is Changing Actors’ Lips to Match Dubbed Voices. Is This the Future of Dubbing?

For decades, dubbing has asked audiences to accept a small but obvious compromise. The voice may sound convincing, the translation may work and the performance may be excellent, but the actor’s mouth on screen is still moving to words originally spoken in another language.

Prime Video is now trying to make that disconnect much harder to see. Its new lip-sync technology uses artificial intelligence and visual effects to synchronize an actor’s mouth movements with human-dubbed dialogue. The technology effectively changes the picture to better match the translated performance.

The technology has made its debut on the English lip-sync dubs of Seasons 1 and 2 of the German series Maxton Hall: The World Between Us. Prime Video says it will also be used when the third and final season arrives on December 9, with plans to expand the technology to additional titles in the future.

For an industry already having difficult conversations about artificial intelligence and performance, there is an important detail here: the dubbed voices are still human. In this particular use of AI, it is the mouth on screen that is being technologically adjusted, not the voice actor being replaced.

Prime Video Has a New Answer to an Old Dubbing Problem

Anyone who has watched dubbed television has probably noticed the problem Prime Video is trying to solve. A translated sentence can convey exactly the same idea as the original dialogue while requiring different words, different sounds and sometimes a different amount of time to say.

Traditional dubbing has therefore required a careful balancing act. Translators, adaptation writers, directors and performers have to preserve meaning and emotion while also working within the timing and visual constraints of a performance that has already been filmed.

Perfect synchronization is difficult because languages do not behave identically. The mouth shapes produced by particular consonants and vowels vary, while a short expression in one language may require a noticeably longer sentence in another.

Prime Video’s approach changes part of that equation. Instead of requiring the translated performance to do all of the adapting, the image itself can now be altered so the actor’s mouth more closely corresponds with the human-dubbed audio.

Amazon says AI and VFX technologies are being used behind the scenes to power the process. The company describes the aim as creating a more seamless viewing experience while reducing the visual disconnect that can occur with traditional dubbing.

That might sound like a relatively minor technical improvement. If it works convincingly across different productions, however, it could change what audiences expect when they select a dubbed audio track.

The Voice Is Still Human

The distinction between AI-assisted lip synchronization and an AI-generated voice is particularly important for the voice-over industry. Prime Video’s announcement specifically describes the system as synchronizing actors’ lip movements with human-dubbed audio, rather than generating replacement dialogue with synthetic voices.

That means the emotional interpretation of the translated dialogue still comes from a human dubbing performer. The actor recording the localized version still has to interpret the character, respond to the scene and make decisions about pace, emotion, emphasis, humour, tension and dozens of other details that contribute to a believable performance.

What has changed is the relationship between that performance and the image. Historically, the picture has been fixed, leaving the translated script and voice performance to work around movements captured when the scene was originally filmed.

Visual lip-sync technology introduces the possibility that both sides can adapt. The dubbed performance still has to work with the timing and dramatic rhythm of the scene, but the visible mouth movements can potentially move closer to the translated dialogue rather than forcing every compromise onto the localization team.

That could have interesting consequences for dubbing actors and directors. If a translated line does not have to conform quite as rigidly to every visible mouth movement, there may eventually be more room to prioritize natural language and performance, although Prime Video has not said that its new technology will change how dubbing sessions themselves are directed.

For now, that possibility remains an inference rather than an announced change to the production process. What Amazon has confirmed is that the technology is being applied to human-dubbed performances and that the company intends to use it on more titles.

Why Start With Maxton Hall?

Prime Video has not chosen an obscure catalogue title for the debut. Amazon describes Maxton Hall as its most-watched International Original series ever, making the German romantic drama a particularly visible test of whether enhanced synchronization can make dubbed entertainment feel more natural to international viewers.

Prime Video has also created a separately identified version of the programme called Maxton Hall: The World Between Us (with Lip Sync for English Dub). That version is now listed alongside the original on Prime Video, giving the technology a much more public debut than a behind-the-scenes production experiment.

The choice makes sense in another way. Streaming has made television produced in one country immediately accessible to viewers thousands of miles away, and the audience encountering an international series may have no particular attachment to either subtitles or dubbing.

Some viewers want to hear the original actors and read subtitles, while others prefer to concentrate on the picture and listen in their own language. A more convincing visual match could make dubbing more appealing to viewers who previously found the mismatch between voices and mouth movements distracting.

That is ultimately where the success or failure of the technology may be decided. The impressive part will not necessarily be making viewers admire the lip-sync technology, but making them stop thinking about lip-sync altogether.

Amazon Has Experimented With AI and Dubbing Before

This is not Prime Video’s first experiment with artificial intelligence in localization, but it is important not to confuse the new lip-sync feature with the company’s earlier work.

In March 2025, Amazon announced an AI-aided dubbing pilot involving 12 licensed films and series that did not already have dubbing support. Amazon described that project as a hybrid approach involving localization professionals working alongside AI, with human professionals providing quality control.

The Maxton Hall announcement concerns something different. Amazon is explicitly describing the new system as visual dubbing in which AI and VFX synchronize the actors’ mouth movements to audio that has been dubbed by humans.

For voice actors, that difference is significant. Discussion of “AI dubbing” can encompass several very different technologies, from synthetic speech and cloned voices to translation tools, editing systems and now the digital alteration of an actor’s lips.

Treating all of those developments as the same thing risks obscuring where human performers remain involved and where automation may actually be changing their work. Prime Video’s new feature is an example of AI operating around a human voice performance rather than simply generating one.

What Could This Mean for Dubbing Actors?

One of the more interesting possibilities is that better visual synchronization could eventually give localized performances greater flexibility. Traditional dubbing sometimes requires wording to be shortened, expanded or altered partly because a translated line must fit the timing and visible movements of the original actor.

If technology can reduce some of that visual mismatch after the voice has been recorded, the balance may begin to shift. Adaptation writers and performers could potentially have more room to choose dialogue that sounds natural in the target language instead of chasing every visible mouth shape.

There are limits to that idea. A dubbed performance still has to fit the scene, interact with other characters and respect pauses, cuts, reactions and dramatic timing, so visual technology cannot simply remove all of the constraints involved in dubbing.

There are also creative questions that extend beyond voice acting. An actor’s mouth movement is part of a screen performance, and altering it after filming means changing something the original performer physically did on camera.

Amazon says its technology is being used under creative oversight to preserve the integrity of the original artistic vision. As the technology expands to more productions, exactly how that oversight operates may become an important part of the conversation.

A subtle adjustment that makes translated dialogue feel natural may be welcomed very differently from an alteration that viewers perceive as changing an actor’s expression or performance. The technology therefore has to succeed not only technically, but creatively.

If It Works, Will Audiences Even Notice?

There is an unusual paradox at the centre of this experiment. The better Prime Video’s technology becomes, the less viewers may notice that it exists.

A distracting dub makes itself obvious whenever the spoken dialogue and visible mouth movements drift apart. A convincing dub, by contrast, allows the viewer to concentrate on the characters and story rather than the mechanics of localization.

Prime Video appears to be betting that altering the picture can help close that gap. The company says the goal is a more seamless and immersive experience for global audiences, while also helping creators reach viewers across different languages.

The technology is also arriving at a moment when international entertainment is easier to discover than ever through global streaming platforms. As audiences become accustomed to moving between languages, the quality of localization becomes part of the experience rather than something that happens invisibly after production.

That makes the Maxton Hall experiment worth watching beyond the novelty of AI-adjusted lips. If viewers respond positively and Prime Video follows through on its plan to introduce the technology on more titles, visually enhanced dubbing could gradually become another expected option on major streaming releases.

It is still far too early to declare that the traditional dub is disappearing. Prime Video has demonstrated the technology on one major international series, and there is not yet enough evidence to know whether audiences will prefer it consistently across genres, languages and performances.

What the announcement does show is that one of dubbing’s oldest limitations is no longer necessarily fixed. For generations of dubbing professionals, the translated performance had to adapt to a picture that could not change, but now the picture can adapt too.

For voice actors, that makes this development more complicated than another story about AI replacing human performers. The human voice remains at the centre of this particular dub, while artificial intelligence is changing what happens around that performance.

And that may be the bigger story to watch. As AI becomes embedded in more parts of entertainment production, the question may increasingly be not simply whether a performance is human or artificial, but how much technology has reshaped what audiences ultimately see and hear.

Would more convincing lip movements make you more likely to watch a dubbed film or series, or would you rather the original screen performance remain untouched?

Photo by Manuel Luikenga on Unsplash

VoiceEditSuite.com

Clean a session in ninety seconds, not three hours.

Pro tools built for working voice actors. Upload a raw take, get back a clean one.

$20/month
$10/month
Launch price until Sep 30
Auto Clean Up ACX Audiobook Prep Audition Director Performance Coach FREE Studio Analysis FREE Rate Calculator
Cancel anytime. No contract. SHOW ME MORE