AI Audio · text to speech

Text to Speech

Text to Speech은 “AI text to speech” 검색을 위한 LipSyncing AI 페이지입니다. 홈을 베낀 글이 아니라 스튜디오가 이 작업 모드로 열립니다. 왼쪽 메뉴의 AI Audio에 있습니다. 파일을 올리고 대본이나 오디오를 더한 뒤 생성하세요.

Text to speech

립싱크 생성

150 credits · $1.50

Pick a voice from the library, paste a script, then generate. My voices are clones and designs you saved. Public voice AI is the catalog for this engine.

My voices0

No saved voices for this engine yet. They appear here after you clone or design one. You can still pick a public voice below.

Public voice AI10

Staff-curated public catalog for this engine, plus live voices when a key is pasted. Play a sample here. Open Text to Speech when you want a generated track.

Showing the built-in catalog. Staff can edit it on Admin → Voice Library.

Text to speech runs on ElevenLabs, Cartesia, and Chatterbox (RunPod). Pick a voice from My voices or public voice AI. Paste keys on Admin → Models.

See an example

InfiniteTalk example from Wavespeed. Sample generated by this tool’s model. Your clip uses your photo, script, or prompt.
Voice sample ready to clone
Voice sample ready to clone
Script turned into speech
Script turned into speech
Audio attached to a talking avatar next
Audio attached to a talking avatar next

이 text to speech 페이지가 있는 이유

My voices, then public AI

Clones and designed voices you saved sit first. Public catalog voices sit second. Pick one and generate.

Emotion without a second take

Calm, bright, or urgent — match the landing page, not a flat read.

Made to hand off

The file is an input for talking photo, talking avatar, dubbing, and singing (spoken intros).

Script-length honesty

TTS is cheap compared with a crew. The video job still bills the 30-second block plus extras.

사용 방법

  1. Step 1: record or upload a sample, or pick a voice
    1단계

    Pick a voice from the library

    My voices are clones and designs you already saved. Public voice AI is the catalog.

  2. Step 2: paste the script
    2단계

    Paste the script and generate

    Listen before you spend a talking-video job.

  3. Step 3: generate speech for lip sync
    3단계

    Send to a face tool

    Talking avatar, photo lip sync, or video lip sync.

What this clip costs in credits

Every text to speech job uses the same meter: 150 credits ($1.50) for the first 30 seconds, then 5 credits ($0.05) per extra second. Subscriptions price credits at $0.01; one-time packs are $0.08 each.

DurationcreditsGeneration price
15s150$1.50
30s150$1.50
45s225$2.25
60s300$3.00
90s450$4.50
120s600$6.00

First 30 seconds = 150 credits ($1.50). Plan credits roll over while you stay subscribed.

어디에 쓰이나

Avatar reads

Generate audio, then open talking avatar on the same script.

Dubbing drafts

TTS in the target language before you hire a human mixer.

IVR-adjacent help clips

Consistent library voice across 40 macros.

text to speech 질문

관련 키워드 페이지

각각 다른 검색을 노립니다. 왼쪽 메뉴로 같은 클러스터에 머무르세요.