Home
Assets
My Works
Voice Library
Developer Docs
Blog
English
Pricing
Sign In
Lip Sync
HomeVideo ToolsLip Sync

Video Tools · PixVerse

Lip Sync

Give the person on screen a new line

HOW IT WORKS

How to lip sync a video

Upload a clip of someone talking and the audio you want them to say. Their mouth is matched to the voice track you supplied.

  1. 01

    Upload the person clip

    A video of the speaker, up to 60 seconds and 100 MB. A clear, front-facing face gives the best match.

  2. 02

    Upload the voice track

    An MP3 or WAV between 5 and 60 seconds, up to 100 MB. This is the audio the mouth will follow.

  3. 03

    Check the source

    The finished clip inherits the quality, ratio and length of the video you uploaded, so this page keeps those controls out of the way.

  4. 04

    Generate and download

    Check the credit cost, preview the sync, then download the result.

Q&A

AI lip sync FAQ

01

Which model powers this page?

PixVerse, which is the one model offering lip sync, so the page shows a fixed model rather than a list.

02

How do I supply the voice?

Upload an audio file — MP3 or WAV, between 5 and 60 seconds. The mouth movements are matched to that recording.

03

How long can the person clip be?

Up to 60 seconds, at 100 MB or less. The generated video keeps the length of the clip you upload.

04

Can I set the quality and aspect ratio?

The result inherits both from your source clip, so those controls stay off this page and what you upload determines what you get back.

05

What footage works best?

A steady shot where the face is clearly visible and facing the camera, with the mouth unobstructed.

06

How many credits does it cost?

It depends on the length of your audio and source clip, and the figure appears beside the generate button before you confirm.

What lip sync is good for
07

Another language, same presenter

Supply a translated voice track and keep the person on screen the same.

08

Fixing a line after the shoot

Re-record one sentence and match it back to the footage instead of reshooting.

09

Spokesperson variants

Produce several versions of a message from one piece of footage.

10

Course and training updates

Refresh the narration of an existing module while keeping the original presenter.

ListenHub

One platform for AI content creation — from an idea to audio, images, slides and video.

Audio
AI PodcastText to SpeechMulti-Speaker VoiceoverAudio to TextVoice CloningAI Voice
Video
AI VideoExplainer VideoLip SyncAd GeneratorPromo VideoMotion Transfer
Image
AI ImageSlides
Resources
PricingDiscoverVoice LibraryBlogAPI DocumentationMCP Server GuideAgent Skills Guide
Company
About UsContact UsTerms of UsePrivacy Policy
© 2026 MarsWave

Needs Source video + Voice audio

Source videoRequired
Voice audioRequired
ⓘQuality, aspect ratio and length all follow the clip you upload.