Vibes AI lip sync

Create AI Lip-Sync Videos

Lip sync makes a face speak or sing words it never said, by re-rendering the mouth to match new audio. Meta lists lip sync among Vibes features; VibesAI has a dedicated composer for it that accepts a video plus uploaded or generated audio. For a still portrait, the talking avatar path does the same job from one photo. The composer below is the lip-sync tool.

Last reviewed September 20, 2026

Independent AI creation platform. Not affiliated with or endorsed by Meta or Meta Vibes.

Video

Add the video

A talking clip up to 60s

New audio

Add voice audio

MP3, WAV or M4A up to 50MB

Sign in to see your credit estimate. Creating needs a plan, from $4.99 a week.

Independent AI creation platform. Not affiliated with or endorsed by Meta or Meta Vibes.

What AI lip sync does

Given a face on video and an audio track, the model re-renders the mouth, jaw and nearby skin frame by frame so the lips match the phonemes in the audio. The rest of the performance is untouched.

Photo, character or video inputs

InputToolResult
A still portraitTalking avatarPresenter with synced lips, blinks and head motion
A saved characterTalking avatar with the character attachedYour character speaking, matching your other videos
An existing videoLip syncThe original performance with a new mouth track

Dialogue versus song

Dialogue

Clean speech syncs most accurately. Translation and line fixes are the common uses.

Songs

Vocals sync well when isolated. You need rights to the recording.

Voice generation

No recording? Type a script, choose a voice, sync it.

Best input practices

  • Front-facing, steady footage with the mouth visible.
  • Audio without background music for speech.
  • Match audio and clip length or trim in the composer.
  • Keep clips under a minute; chain for longer pieces.

Consent and copyright

Use your own voice and footage, or material you are licensed to use. Impersonating real people is against the guidelines.

Lip-sync and voice models

Live from our model registry. Availability and specs update as models change.

Lip Sync

Lip-sync an existing video to new audio.

audio

Frequently asked

Can I make a photo talk?
Yes, with the talking avatar: one portrait plus a script or audio file.
Can I translate a video I made?
Yes. Generate or record the translated speech, then lip-sync the original clip to it.
Does it work on animated characters?
Illustrated and generated faces work where the routed model supports non-photoreal faces; results vary with how human-like the face is.
How much does a generation cost?
Every generation is quoted in credits before you confirm, based on the routed model, duration and resolution. Credits are reserved when you start and released automatically if the generation fails, so you only pay for what completes. Plans start at $4.99 a week and include a credit allowance; subscribers can top up with credit packs.
Is this website Meta Vibes?
No. VibesAI is an independent AI creation platform and is not affiliated with, sponsored by or endorsed by Meta. We describe Meta Vibes on these pages to answer the questions people search for, and we offer our own tools as an alternative.

Sources and update note

Third-party facts are reviewed on the cadence shown by the last-reviewed date; if you spot something outdated, tell us and we will correct it.

Ready when you are.

Plans from $4.99 a week. Publish when you're proud of it.

Lip-sync a video