> ## Documentation Index
> Fetch the complete documentation index at: https://support.myapps.ai/llms.txt
> Use this file to discover all available pages before exploring further.

# Audio FAQ

> Common questions about music, voice, and sound in Pixio

## Choosing a model

<AccordionGroup>
  <Accordion title="I want a song">
    [SongCraft](/families/songcraft) generates full songs from a prompt. [MiniMax](/families/minimax) has notable audio depth too, and [Mureka](/families/audio-avatar-roundup) is worth comparing.
  </Accordion>

  <Accordion title="I want speech or narration">
    [ElevenLabs](/families/elevenlabs) is the default for text-to-speech and voice work.
  </Accordion>

  <Accordion title="I want to change an existing voice or audio track">
    Filter for **audio-to-audio**. [ElevenLabs](/families/elevenlabs) and [MiniMax](/families/minimax) both offer it.
  </Accordion>

  <Accordion title="I want sound effects rather than music">
    Look at **text-to-audio** models — they cover effects as well as music. Read the model descriptions, since capability names don't distinguish the two.
  </Accordion>

  <Accordion title="How do I pick between them?">
    Audio quality is subjective in a way image quality isn't. Generate the same prompt on two or three models and listen. There's no substitute for that.
  </Accordion>
</AccordionGroup>

## Voice and video together

<AccordionGroup>
  <Accordion title="How do I make a character speak?">
    Two steps. Generate the voice track (**Generate Voice** [chained action](/generate#chained-actions) or [ElevenLabs](/families/elevenlabs)), then sync it — **Add Lipsync** on an image, **Lipsync Video** on a clip, or [Sync](/families/sync) directly.
  </Accordion>

  <Accordion title="Can I drive video from audio?">
    Yes — **audio-to-video** models generate video from an audio track. [Kling](/families/kling) and [LTX](/families/ltx) both have them.
  </Accordion>

  <Accordion title="I need a presenter reading a script">
    [HeyGen](/families/heygen) handles presenter video directly, rather than assembling voice and lipsync yourself.
  </Accordion>
</AccordionGroup>

## Costs

<AccordionGroup>
  <Accordion title="How is audio billed?">
    Often **per second**, like video — so a long track costs proportionally more. The cost is shown before you submit. See [Credits & Plans](/credits).
  </Accordion>

  <Accordion title="How do I test cheaply?">
    Generate a short clip to check the voice or musical style before committing to full length. Style is audible in a few seconds.
  </Accordion>
</AccordionGroup>

## Practical

<AccordionGroup>
  <Accordion title="Where do my audio files go?">
    [My Assets](/assets) → Generated, alongside images and video. Download from there or from the preview.
  </Accordion>

  <Accordion title="Can I use audio I already have?">
    Yes — upload it in the form or pick it from your assets, for lipsync, audio-to-video, or audio transformation.
  </Accordion>

  <Accordion title="Can I put audio on a video timeline?">
    Yes — the [Video Editor](/video-editor) assembles clips and audio. Generate the pieces, then arrange them there.
  </Accordion>

  <Accordion title="Audio generation is failing">
    Check [status](/status) — audio is its own service group. Failed runs aren't charged.
  </Accordion>
</AccordionGroup>

## Related

* [ElevenLabs](/families/elevenlabs) · [SongCraft](/families/songcraft) · [MiniMax](/families/minimax) · [Sync](/families/sync)
* [Generate](/generate) · [Credits & Plans](/credits)
