AI Video · reference to video
Reference to Video
Reference to video is the multimodal slot. When a prompt is not precise enough, you @-mention stills, a motion clip, or a sound bed so the model keeps the parts you already like. It is how cinematic generators on this site take four inputs at once. It is not motion control (which copies a performance 1:1 onto a character still) and not ad clone (which copies an ad framework onto a SKU).
Make a clip
Generate from references
150 credits · $1.50
Each model has its own length and resolution caps. Seconds and quality chips update when you switch Seedance 2.0, Seedance 2.5, Veo 3.1, or Wan 2.5. Paste a Wavespeed key on Admin → Models.
See an example
Why people use reference to video
Multiple references
Face from still A, room from still B, camera from a clip — say what each reference is for.
Audio as a reference
Some models can follow a bed or a line. Lip sync still wins when words must land.
Character consistency
Reuse the same avatar still as a reference instead of hoping the prompt remembers them.
Model switch in one studio
Try a draft model, then a heavier one, on the same reference set.
Three simple steps
- Step 1
Add the references
Image, video, and optional audio. Label them in the prompt.
- Step 2
Write what should change
Keep identity, change location — be explicit.
- Step 3
Generate, then mouth-pass if it speaks
Video lip sync or talking avatar on the take.
What this clip costs in credits
Every reference to video job uses the same meter: 150 credits ($1.50) for the first 30 seconds, then 5 credits ($0.05) per extra second. Subscriptions price credits at $0.01; one-time packs are $0.08 each.
| Duration | credits | Generation price |
|---|---|---|
| 15s | 150 | $1.50 |
| 30s | 150 | $1.50 |
| 45s | 225 | $2.25 |
| 60s | 300 | $3.00 |
| 90s | 450 | $4.50 |
| 120s | 600 | $6.00 |
First 30 seconds = 150 credits ($1.50). Plan credits roll over while you stay subscribed.
Good for
Short films and spots
Lock a character, vary the scene.
Product stories
Pack still + lifestyle clip + VO.
Social series
Same face reference every episode.
Common questions about reference to video
You may also like
Each of these is a different job. Use the left menu, or tap one of these cards.
AI Video Generator
The LipSyncing AI video generator creates clips from text, images, or reference media, then you can run a lip sync pass on any talking result.
Motion Control
AI motion control transfers movement from a reference video onto a character image. Keep identity, copy the dance, gesture, or talking stance.
Frames to Video
Provide a first and last frame and generate the motion between them. Use frames-to-video when you already know how the shot should open and close.
Image to Video
Image to video AI turns a photograph or illustration into a moving clip. Use it to create a talking-ready shot from a single frame.