Create AI Lip-Sync Videos
Lip sync makes a face speak or sing words it never said, by re-rendering the mouth to match new audio. Meta lists lip sync among Vibes features; VibesAI has a dedicated composer for it that accepts a video plus uploaded or generated audio. For a still portrait, the talking avatar path does the same job from one photo. The composer below is the lip-sync tool.
Last reviewed September 20, 2026
Independent AI creation platform. Not affiliated with or endorsed by Meta or Meta Vibes.
Add the video
A talking clip up to 60s
Add voice audio
MP3, WAV or M4A up to 50MB
Sign in to see your credit estimate. Creating needs a plan, from $4.99 a week.
Independent AI creation platform. Not affiliated with or endorsed by Meta or Meta Vibes.
What AI lip sync does
Given a face on video and an audio track, the model re-renders the mouth, jaw and nearby skin frame by frame so the lips match the phonemes in the audio. The rest of the performance is untouched.
Photo, character or video inputs
| Input | Tool | Result |
|---|---|---|
| A still portrait | Talking avatar | Presenter with synced lips, blinks and head motion |
| A saved character | Talking avatar with the character attached | Your character speaking, matching your other videos |
| An existing video | Lip sync | The original performance with a new mouth track |
Dialogue versus song
Dialogue
Clean speech syncs most accurately. Translation and line fixes are the common uses.
Songs
Vocals sync well when isolated. You need rights to the recording.
Voice generation
No recording? Type a script, choose a voice, sync it.
Best input practices
- Front-facing, steady footage with the mouth visible.
- Audio without background music for speech.
- Match audio and clip length or trim in the composer.
- Keep clips under a minute; chain for longer pieces.
Consent and copyright
Use your own voice and footage, or material you are licensed to use. Impersonating real people is against the guidelines.
Lip-sync and voice models
Live from our model registry. Availability and specs update as models change.
Lip Sync
Lip-sync an existing video to new audio.
Frequently asked
- Can I make a photo talk?
- Yes, with the talking avatar: one portrait plus a script or audio file.
- Can I translate a video I made?
- Yes. Generate or record the translated speech, then lip-sync the original clip to it.
- Does it work on animated characters?
- Illustrated and generated faces work where the routed model supports non-photoreal faces; results vary with how human-like the face is.
- How much does a generation cost?
- Every generation is quoted in credits before you confirm, based on the routed model, duration and resolution. Credits are reserved when you start and released automatically if the generation fails, so you only pay for what completes. Plans start at $4.99 a week and include a credit allowance; subscribers can top up with credit packs.
- Is this website Meta Vibes?
- No. VibesAI is an independent AI creation platform and is not affiliated with, sponsored by or endorsed by Meta. We describe Meta Vibes on these pages to answer the questions people search for, and we offer our own tools as an alternative.
Sources and update note
- Meta newsroom and official Vibes product pages · checked September 20, 2026
- Official Meta AI app store listings · checked September 20, 2026
Third-party facts are reviewed on the cadence shown by the last-reviewed date; if you spot something outdated, tell us and we will correct it.
Related
Ready when you are.
Plans from $4.99 a week. Publish when you're proud of it.