Video Tools · PixVerse
Give the person on screen a new line
Upload a clip of someone talking and the audio you want them to say. Their mouth is matched to the voice track you supplied.
A video of the speaker, up to 60 seconds and 100 MB. A clear, front-facing face gives the best match.
An MP3 or WAV between 5 and 60 seconds, up to 100 MB. This is the audio the mouth will follow.
The finished clip inherits the quality, ratio and length of the video you uploaded, so this page keeps those controls out of the way.
Check the credit cost, preview the sync, then download the result.
PixVerse, which is the one model offering lip sync, so the page shows a fixed model rather than a list.
Upload an audio file — MP3 or WAV, between 5 and 60 seconds. The mouth movements are matched to that recording.
Up to 60 seconds, at 100 MB or less. The generated video keeps the length of the clip you upload.
The result inherits both from your source clip, so those controls stay off this page and what you upload determines what you get back.
A steady shot where the face is clearly visible and facing the camera, with the mouth unobstructed.
It depends on the length of your audio and source clip, and the figure appears beside the generate button before you confirm.
Supply a translated voice track and keep the person on screen the same.
Re-record one sentence and match it back to the footage instead of reshooting.
Produce several versions of a message from one piece of footage.
Refresh the narration of an existing module while keeping the original presenter.
Needs Source video + Voice audio