From a single still · photo lip sync
Photo Lip Sync
Photo lip sync is the shortest path from a JPEG to a talking clip. The model finds the face, estimates a mouth cavity, and drives visemes while adding the small blinks and head drift that stop a still from looking like a cardboard cutout. It is the right tool when you have a brand portrait, a team headshot, or a character sheet — and no camera time.
Three steps. That is the whole job.
- 1
Upload a video or photo
A clear, front-facing face works best. Or pick a sample in the studio.
- 2
Add audio or a script
Type the line and pick a voice, drop a file, or record. No timeline.
- 3
Generate lip sync
AI matches the mouth to every word. Then download the MP4.
Make a video
Sync this photo
150 credits · $1.50
Upload a video or photo
Click to add a photo, or drop a file here
Character library
Pick a saved look or a public character, then generate. You can still upload your own still.
No avatar selected yet
Public library
52 looks
Add audio or a script
Voice library
Same catalog as Voice Library: your clones first, then public voice AI for the engine you pick.
Rachel
No clones for this engine yet. Clone or design one, or pick public voice AI. Clone a voice
Public voice AI
10 voices
Generate lip sync
150 credits · $1.50 · Log in
Press Make a video. If we do not have an API key yet, you still get a preview on this computer. AI by Zoice Avatar X
Your video
Nothing here yet. Add a face, add words, then press Make a video. The clip will play in this box.
See an example
Why people use photo lip sync
Built for headshots
Passport-style, LinkedIn, and campaign portraits all work. Chin-up, mouth visible, no heavy shadow across the lips.
Motion that stays modest
The model adds micro-movement, not a dance. Photo lip sync should look like a person talking at a desk, not a deepfake skit.
Script-first workflow
Most photo jobs start with text. Choose a voice, adjust pace, generate. Upload audio only when you already have a read.
Same engine as talking photo
Photo lip sync is the technical name for the talking-photo pipeline. Use this page when you are searching for the image-plus-audio job.
Three simple steps
- Step 1
Upload a clear face
JPG, PNG, or WebP. One subject, front or three-quarter, teeth not hidden by a mask or hand.
- Step 2
Write or attach speech
A few sentences is enough. Longer reads still work; they just cost more credits.
- Step 3
Download the talking still
Export MP4. If the jaw feels stiff, switch from LipSync Fast to Talking 4.0 and regenerate.
What this clip costs in credits
Every photo lip sync job uses the same meter: 150 credits ($1.50) for the first 30 seconds, then 5 credits ($0.05) per extra second. Subscriptions price credits at $0.01; one-time packs are $0.08 each.
| Duration | credits | Generation price |
|---|---|---|
| 15s | 150 | $1.50 |
| 30s | 150 | $1.50 |
| 45s | 225 | $2.25 |
| 60s | 300 | $3.00 |
| 90s | 450 | $4.50 |
| 120s | 600 | $6.00 |
First 30 seconds = 150 credits ($1.50). Plan credits roll over while you stay subscribed.
Good for
Team pages that speak
Turn staff photos into 15-second intros for a careers site or conference booth loop.
Personal messages
A birthday still plus a recorded greeting beats a text card, and you never film yourself.
Casting tests
Hear a voice on a concept portrait before you hire a performer.
Common questions about photo lip sync
You may also like
Each of these is a different job. Use the left menu, or tap one of these cards.
Talking Photo
AI talking photo generator: upload a portrait, add a script or recording, and LipSyncing AI animates the mouth into a speaking video.
Make a Photo Talk
Make a photo talk online with LipSyncing AI. Upload the image, type what it should say, and download a lip-synced talking clip.
Animate Face
Animate a still face to speech with lip sync. Upload a portrait, add a script or recording, and LipSyncing AI rebuilds the mouth on the original pixels.
Talking Avatar
Build a talking avatar from a photo or a library character. Drive it with a script or audio and reuse the same host across every video.