From a prompt · text to video AI
Text to Video
Text to video is for the shot you do not have. You describe the subject, the camera, the light, and the length. The model returns a clip. If that clip is a person facing the lens, it can go straight into lip sync. If it is a city at dawn, it is B-roll. Either way you stayed in one workspace instead of exporting from a toy and importing into a lip sync app.
Make a clip
Generate from text
150 credits · $1.50
Each model has its own length and resolution caps. Seconds and quality chips update when you switch Seedance 2.0, Seedance 2.5, Veo 3.1, or Wan 2.5. Paste a Wavespeed key on Admin → Models.
See an example
Why people use text to video AI
Prompt like a shot list
Subject, action, lens, light. Poetry is optional; specifics render better.
Iterate in place
Change one clause, generate again. Keep the takes you like.
Speech is a second step
Text to video invents picture. Lip sync invents a believable mouth on that picture.
Short by design
These models are clip tools, not feature tools. Stack shots in an editor.
Three simple steps
- Step 1
Write the shot
One sentence for the picture, one for the camera.
- Step 2
Generate a draft
Pick the take with a readable face if you plan to add speech.
- Step 3
Attach a line
Open the lip sync generator on that take.
What this clip costs in credits
Every text to video AI job uses the same meter: 150 credits ($1.50) for the first 30 seconds, then 5 credits ($0.05) per extra second. Subscriptions price credits at $0.01; one-time packs are $0.08 each.
| Duration | credits | Generation price |
|---|---|---|
| 15s | 150 | $1.50 |
| 30s | 150 | $1.50 |
| 45s | 225 | $2.25 |
| 60s | 300 | $3.00 |
| 90s | 450 | $4.50 |
| 120s | 600 | $6.00 |
First 30 seconds = 150 credits ($1.50). Plan credits roll over while you stay subscribed.
Good for
Pitch decks
A moving concept before anyone spends on a crew.
Channel bumpers
A signature shot you can regenerate when the season changes.
Talking characters that do not exist yet
Generate the face in motion, then run a script.
Common questions about text to video AI
You may also like
Each of these is a different job. Use the left menu, or tap one of these cards.
AI Video Generator
The LipSyncing AI video generator creates clips from text, images, or reference media, then you can run a lip sync pass on any talking result.
Image to Video
Image to video AI turns a photograph or illustration into a moving clip. Use it to create a talking-ready shot from a single frame.
Talking Head Video
Generate a talking head video from a photo or avatar, a script, and a voice. Lip-synced, consistent, and ready for YouTube, courses, and ads.
Talking Avatar
Build a talking avatar from a photo or a library character. Drive it with a script or audio and reuse the same host across every video.