> ## Documentation Index
> Fetch the complete documentation index at: https://docs.comfy.org/llms.txt
> Use this file to discover all available pages before exploring further.

# Seed Audio 1.0 Workflows in ComfyUI

> How to run the Seed Audio 1.0 text-to-audio, voice cloning, and image-driven audio workflows in ComfyUI, locally or on Comfy Cloud

This guide shows how to run the Seed Audio 1.0 workflows in ComfyUI. For what the model is and what it does best, see the [Seed Audio 1.0 overview](/tutorials/partner-nodes/bytedance/seed-audio-1-0).

<Tip>
  To use the Partner Nodes, you need to ensure that you are logged in properly and using a permitted network environment. Please refer to the [Partner Nodes Overview](/tutorials/partner-nodes/overview) section of the documentation to understand the specific requirements for using the Partner Nodes.
</Tip>

<Tip>
  <Tabs>
    <Tab title="Local users">
      Make sure your ComfyUI is updated.

      * [Download ComfyUI](https://www.comfy.org/download)
      * [Update Guide](/installation/update_comfyui)

      Workflows in this guide can be found in the [Workflow Templates](/interface/features/template).
      If you can't find them in the template, your ComfyUI may be outdated.

      If nodes are missing when loading a workflow, possible reasons:

      1. You are not using the latest ComfyUI version (Nightly version)
      2. Some nodes failed to import at startup
    </Tab>

    <Tab title="Cloud users">
      * [Cloud](https://cloud.comfy.org) will update after ComfyUI stable release.

      So, if you find any core node missing in this document, it might be because the new core nodes have not yet been released in the latest stable version. Please wait for the next stable release.
    </Tab>
  </Tabs>
</Tip>

## Available workflows

<h3 id="api_bytedance_seed_audio1_0_t2a">
  Seed Audio 1.0: Text to Audio
</h3>

Input a text prompt to generate speech, dialogue, background music, and sound effects in one audio file.

<img src="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/templates/api_bytedance_seed_audio1_0_t2a-1.webp" alt="Seed Audio 1.0: Text to Audio workflow preview" />

<CardGroup cols={2}>
  <Card title="Run on Comfy Cloud" icon="cloud" href="https://cloud.comfy.org/?template=api_bytedance_seed_audio1_0_t2a&utm_source=docs&utm_medium=referral&utm_campaign=seed-audio-1-0">
    Open in Comfy Cloud
  </Card>

  <Card title="Download Workflow" icon="download" href="https://github.com/Comfy-Org/workflow_templates/blob/main/templates/api_bytedance_seed_audio1_0_t2a.json">
    Download JSON or search "Seed Audio 1.0: Text to Audio" in Template Library
  </Card>
</CardGroup>

<h3 id="api_bytedance_seed_audio1_0_ta2a">
  Seed Audio 1.0: Text + Audio to Audio
</h3>

Upload a reference audio clip and write a text prompt to clone the voice into a new scene.

<img src="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/templates/api_bytedance_seed_audio1_0_ta2a-1.webp" alt="Seed Audio 1.0: Text + Audio to Audio workflow preview" />

<CardGroup cols={2}>
  <Card title="Run on Comfy Cloud" icon="cloud" href="https://cloud.comfy.org/?template=api_bytedance_seed_audio1_0_ta2a&utm_source=docs&utm_medium=referral&utm_campaign=seed-audio-1-0">
    Open in Comfy Cloud
  </Card>

  <Card title="Download Workflow" icon="download" href="https://github.com/Comfy-Org/workflow_templates/blob/main/templates/api_bytedance_seed_audio1_0_ta2a.json">
    Download JSON or search "Seed Audio 1.0: Text + Audio to Audio" in Template Library
  </Card>
</CardGroup>

**Input materials**

Upload this file to the matching `LoadAudio` node:

<CardGroup cols={2}>
  <Card title="seed_audio_ref_audio1.mp3" icon="music" href="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/input/seed_audio_ref_audio1.mp3">
    `LoadAudio` node 8 · `seed_audio_ref_audio1.mp3`
  </Card>
</CardGroup>

<h3 id="api_bytedance_seed_audio1_0_ti2a">
  Seed Audio 1.0: Text + Image to Audio
</h3>

Upload a character image and write a text prompt to generate audio with a voice style matching the visual subject.

<img src="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/templates/api_bytedance_seed_audio1_0_ti2a-1.webp" alt="Seed Audio 1.0: Text + Image to Audio workflow preview" />

<CardGroup cols={2}>
  <Card title="Run on Comfy Cloud" icon="cloud" href="https://cloud.comfy.org/?template=api_bytedance_seed_audio1_0_ti2a&utm_source=docs&utm_medium=referral&utm_campaign=seed-audio-1-0">
    Open in Comfy Cloud
  </Card>

  <Card title="Download Workflow" icon="download" href="https://github.com/Comfy-Org/workflow_templates/blob/main/templates/api_bytedance_seed_audio1_0_ti2a.json">
    Download JSON or search "Seed Audio 1.0: Text + Image to Audio" in Template Library
  </Card>
</CardGroup>

**Input materials**

Upload this file to the matching `LoadImage` node:

<CardGroup cols={2}>
  <Card title="girl_and_fish.png" icon="image" href="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/input/girl_and_fish.png">
    `LoadImage` node 6 · `girl_and_fish.png`
  </Card>
</CardGroup>

<div style={{display: 'grid', gridTemplateColumns: 'repeat(2, minmax(0, 1fr))', gap: '1rem', alignItems: 'start'}}>
  <img src="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/input/girl_and_fish.png" alt="girl_and_fish.png" style={{width: '100%', height: 'auto', objectFit: 'contain'}} />
</div>

## How to use Seed Audio 1.0 in ComfyUI

Seed Audio 1.0 ships as the **ByteDanceSeedAudio** built-in node. You can find it in the node menu under ByteDance.

### Choosing a reference mode

The node offers four reference modes that determine how the generated voice is conditioned:

| Mode                | Description                                                                                                        | When to use                                                  |
| ------------------- | ------------------------------------------------------------------------------------------------------------------ | ------------------------------------------------------------ |
| **Text only**       | Describe everything (voice style, emotion, ambience, and dialogue) in the prompt                                   | Quick generation with no reference material                  |
| **Audio reference** | Connect up to 3 reference audio clips (30s max each) and tag them as `@Audio1`, `@Audio2`, `@Audio3` in the prompt | Cloning specific voices or creating multi-character dialogue |
| **Image reference** | Connect one character image; the model derives a voice from it                                                     | Deriving a voice from a visual character design              |
| **Preset voice**    | Select a built-in TTS 2.0 voice from the dropdown                                                                  | Reliable narration without cloning                           |

### Prompt structure

For best results, structure your prompt to describe the audio scene:

* **Voice and emotion**: Describe the tone, age, gender, accent, or emotional delivery
* **Ambience and background**: Describe the acoustic environment (quiet studio, busy street, concert hall)
* **Sound effects**: Describe specific sounds at the right moment in the timeline
* **Dialogue lines**: Write the lines to speak, naming characters for multi-speaker scenes

**Example: narration with ambience**

> A calm, warm male voice speaks confidently. Soft rain falls in the background, with occasional distant thunder. "Seed Audio 1.0 generates natural speech with full control over pacing, pitch, and presence."

**Example: multi-speaker dialogue with audio references**

> @Audio1 says nervously in a low, fast whisper: "Did you see that?" @Audio2 replies in a calm, slow drawl: "See what? It's just the wind." Light wind ambience throughout.

### Parameter tuning

After connecting your inputs, tune these parameters for the desired output:

* **Speech rate**: Default 0 (normal). Range -50 (0.5x) to 100 (2.0x). Use negative values for relaxed narration, positive for energetic voiceovers
* **Pitch**: Shift in semitones (-12 to +12). Useful for character voices or matching a specific vocal range
* **Loudness**: Default 0 (normal). Range -50 (0.5x) to 100 (2.0x). Adjust for consistent volume across different source materials
* **Sample rate**: Choose from 8kHz to 48kHz. 24kHz is a good default for speech; 44.1kHz or 48kHz for music
* **Seed**: Controls whether the node reruns; results are non-deterministic regardless of seed value

### Edge cases and limitations

* **Reference clip duration**: Audio reference clips are limited to 30 seconds each. For longer reference material, trim or edit the clip before connecting
* **Audio reference ordering**: Reference inputs must be connected sequentially (1, 2, 3) without gaps
* **Tag limit**: A maximum of 3 `@AudioN` tags are supported in a single prompt
* **Character limit**: Prompts are limited to 3000 characters
* **Duration limit**: Maximum 2 minutes of audio per run

## Get started

1. Update ComfyUI to the latest version
2. Add the **ByteDanceSeedAudio** node from the node menu, or open one of the workflow templates above
3. Choose your reference mode and connect inputs
4. Write your prompt describing the audio scene and dialogue
5. Run the workflow
