Lipsync node
Match a person's mouth in a video to a piece of audio.
Lipsync takes a video of a person and a separate audio clip, and redraws the mouth so it matches what is being said. The rest of the shot stays as it was.
It needs both. Wire the clip into the video input and the speech into the audio input. The audio usually comes from a Voice node, but any audio file works.
- Wire your video into the node.
- Wire your audio into the audio input.
- Press Lipsync.
There are three models. Sync 3 is the newest and the default — reach for it on close-ups, profiles, and anything going to a client. Lipsync 2 Pro is stronger on beards and teeth. Lipsync 2 is the cheap one for rough passes. The two Lipsync 2 models add an expressiveness slider. Turn it up if the mouth moves too little, down if it looks overdone. They also have a setting for hands or objects crossing the face. Sync 3 does both of those itself, so neither control appears for it.
The person in your source clip should already be talking. These models redraw an existing mouth movement — they do not start one. Give a still, closed mouth and you will get a small movement at best. When you generate the source clip, ask for someone speaking to camera.
It also works best when the face is clear, reasonably large in frame, and facing roughly towards camera. A face in deep shadow or turned away is hard to match. If there are several people on screen, turn on the active-speaker setting so only the person talking is changed.
Still stuck? Support has the common problems and a way to reach us.