> ## Documentation Index
> Fetch the complete documentation index at: https://docs.comfy.org/llms.txt
> Use this file to discover all available pages before exploring further.

# Vidu Q4 Preview：ComfyUI 图像与参考图生视频工作流

> 使用 ComfyUI 中的 Vidu Q4 Preview 合作节点，从首帧或最多 15 张参考图生成 3 到 16 秒并带原生音频的视频。

**Vidu Q4 Preview** 是 Vidu 推出的视频生成模型，在 ComfyUI 中通过两个合作节点提供。它能生成 3 至 16 秒、带原生音频的片段，画面、对白和音效都由模型在单次生成中一并产出。输出分辨率从 540p 到 4K。

在 ComfyUI 中，**Vidu Q4 Image-to-Video Generation** 可依据单张首帧图像并配合可选提示词生成动画，**Vidu Q4 Reference-to-Video Generation** 则可基于最多 15 张参考图像、可选的参考音频和提示词生成片段。两者都是云端 API 节点：需要带积分的 Comfy 账户，不会下载任何本地模型。每秒计费标准请参见[合作伙伴节点定价](/zh/tutorials/partner-nodes/pricing)。

图生视频会保持输入图像的宽高比。参考生视频需要显式指定 `aspect_ratio`，两个节点都提供 `duration`、`resolution` 以及 `audio` 开关。

## 可用工作流

### 图生视频

为一张图像制作动画。该图像是视频片段的首帧，输出会保持其宽高比，因此当你需要不同的形状时，请裁剪图像。提示词为可选，用于描述镜头中发生的内容。

<img src="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/templates/api_vidu_q4_preview_i2v-1.webp" alt="Vidu Q4 Preview 图生视频工作流预览" />

<CardGroup cols={2}>
  <Card title="在 Comfy Cloud 上运行" icon="cloud" href="https://cloud.comfy.org/?template=api_vidu_q4_preview_i2v&utm_source=docs&utm_medium=referral&utm_campaign=vidu-q4-preview">
    在 Comfy Cloud 中打开
  </Card>

  <Card title="下载工作流" icon="download" href="https://github.com/Comfy-Org/workflow_templates/blob/main/templates/api_vidu_q4_preview_i2v.json">
    下载 JSON，或在模板库中搜索“Vidu Q4 Preview: Image to Video”
  </Card>
</CardGroup>

**输入素材**

<Card title="model_turquoise_yellow_outfit.png" icon="image" href="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/input/model_turquoise_yellow_outfit.png">
  加载到为 Vidu Q4 图生视频节点提供输入的 `LoadImage` 节点中
</Card>

### 参考生视频

使用参考图像、可选的参考音频和提示词构建视频片段。最多可连接 15 张参考图像和 3 段音频，然后在提示词中按顺序引用它们：`image 1`、`image 2`，依此类推。该模板使用两张参考图像，以在整个镜头中保持角色和道具的一致性。

<img src="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/templates/api_vidu_q4_preview_r2v-1.webp" alt="Vidu Q4 Preview 参考生视频工作流预览" />

<CardGroup cols={2}>
  <Card title="在 Comfy Cloud 上运行" icon="cloud" href="https://cloud.comfy.org/?template=api_vidu_q4_preview_r2v&utm_source=docs&utm_medium=referral&utm_campaign=vidu-q4-preview">
    在 Comfy Cloud 中打开
  </Card>

  <Card title="下载工作流" icon="download" href="https://github.com/Comfy-Org/workflow_templates/blob/main/templates/api_vidu_q4_preview_r2v.json">
    下载 JSON，或在模板库中搜索“Vidu Q4 Preview: Reference to Video”
  </Card>
</CardGroup>

**输入素材**

<CardGroup cols={2}>
  <Card title="clay_man_smiley_sweater.png" icon="image" href="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/input/clay_man_smiley_sweater.png">
    加载到第二个参考图像槽位 · `image 2`
  </Card>

  <Card title="yellow_fuzzy_smiley.png" icon="image" href="https://raw.githubusercontent.com/Comfy-Org/workflow_templates/main/input/yellow_fuzzy_smiley.png">
    加载到第一个参考图像槽位 · `image 1`
  </Card>
</CardGroup>

### 工作流总览

两个模板都使用一个小型节点图：

* **LoadImage**：提供首帧（图生视频）或参考图像（参考生视频）
* **Vidu4ImageToVideoNode** / **Vidu4ReferenceVideoNode**：核心节点，配置为使用 `Vidu Q4 Preview` 模型
* **SaveVideo**：写入已完成的视频片段

### 运行步骤

1. **加载输入**：为图生视频设置首帧，或为参考生视频设置参考图像
2. **选择模型**：在节点上保留 `Vidu Q4 Preview`
3. **选择分辨率和时长**：从 `540p` 到 `4K`，时长 3 到 16 秒
4. **切换音频**：保持开启可保留对话和音效，关闭则生成静音视频片段
5. **点击 Queue** 或按 `Ctrl+Enter` 生成

## 节点控件

| 控件 | 取值 | 适用于 | 效果 |
| - | - | - | - |
| `image` | 一张图像 | Image to Video | 视频的首帧。宽高比必须介于 1:5 与 5:1 之间 |
| `reference_images` | 最多 15 张 | Reference to Video | 每个槽位一张图像；单个槽位上的批处理会按图像张数分别计数。在提示词中按顺序以 `image 1`、`image 2` 引用它们 |
| `reference_audios` | 最多 3 个，每个 3 到 12 秒 | Reference to Video | 声音参考。仅使用声音本身，而非其中的词句。需要开启 `audio`。在提示词中指定声音，例如 `image 1 says "Hello!" in the voice from audio 1` |
| `prompt` | 文本，最多 5000 个字符 | 两个节点 | 对 Image to Video 为可选，对 Reference to Video 为必填 |
| `aspect_ratio` | `16:9`、`9:16`、`1:1`、`3:4`、`4:3` | Reference to Video | 输出宽高比。Image to Video 则跟随输入图像 |
| `resolution` | `540p`、`720p`、`1080p`、`2K`、`4K` | 两个节点 | 输出分辨率，默认 `720p` |
| `duration` | 3 到 16 秒 | 两个节点 | 滑块，默认 5 |
| `audio` | on（默认）、off | 两个节点 | 保留原生音频，包括对话和音效 |
| `seed` | 整数 | 两个节点 | 即使使用相同的种子，多次运行的结果仍可能不同 |

## 提示词技巧

* **描述镜头、台词和声音**。模型会同时渲染画面和音频，因此请在同一提示词中写出动作、加引号的台词以及环境声。
* **按顺序标注参考**。参考图像以 `image 1`、`image 2` 的方式引用。请说明每个参考需要保留什么，例如从 `image 1` 中保留角色的脸和服装。
* **为声音参考编写对白**。参考音频承载的是音色而不是词句，所以请在提示词中写出台词，并把它指向某个音色，例如 `image 1 says "Welcome back!" in the voice from audio 1`。
* **让宽高比保持在允许范围内**。参考图像的比例必须介于 1:5 和 5:1 之间。
* **不同运行之间会有差异**。种子不会锁定结果；相同的设置仍可能生成不同的演绎。

## 开始使用

1. 将 ComfyUI 更新到最新版本
2. 前往模板库，搜索 `Vidu Q4 Preview`


This documentation is built and hosted on [Mintlify](https://mintlify.com), a developer documentation platform.