Choosing a model
I want a song
I want a song
I want speech or narration
I want speech or narration
ElevenLabs is the default for text-to-speech and voice work.
I want to change an existing voice or audio track
I want to change an existing voice or audio track
Filter for audio-to-audio. ElevenLabs and MiniMax both offer it.
I want sound effects rather than music
I want sound effects rather than music
Look at text-to-audio models — they cover effects as well as music. Read the model descriptions, since capability names don’t distinguish the two.
How do I pick between them?
How do I pick between them?
Audio quality is subjective in a way image quality isn’t. Generate the same prompt on two or three models and listen. There’s no substitute for that.
Voice and video together
How do I make a character speak?
How do I make a character speak?
Two steps. Generate the voice track (Generate Voice chained action or ElevenLabs), then sync it — Add Lipsync on an image, Lipsync Video on a clip, or Sync directly.
Can I drive video from audio?
Can I drive video from audio?
I need a presenter reading a script
I need a presenter reading a script
HeyGen handles presenter video directly, rather than assembling voice and lipsync yourself.
Costs
How is audio billed?
How is audio billed?
Often per second, like video — so a long track costs proportionally more. The cost is shown before you submit. See Credits & Plans.
How do I test cheaply?
How do I test cheaply?
Generate a short clip to check the voice or musical style before committing to full length. Style is audible in a few seconds.
Practical
Where do my audio files go?
Where do my audio files go?
My Assets → Generated, alongside images and video. Download from there or from the preview.
Can I use audio I already have?
Can I use audio I already have?
Yes — upload it in the form or pick it from your assets, for lipsync, audio-to-video, or audio transformation.
Can I put audio on a video timeline?
Can I put audio on a video timeline?
Yes — the Video Editor assembles clips and audio. Generate the pieces, then arrange them there.
Audio generation is failing
Audio generation is failing
Check status — audio is its own service group. Failed runs aren’t charged.
