Wan2.1 ComfyUI 原生(native)工作流示例
模型安装
本指南中提到的所有模型都可以在这里找到。以下是本指南中的示例所需的常见模型,你可以提前下载: 从文本编码器中选择一个版本下载:umt5_xxl_fp16.safetensors
FP16 精度文本编码器。放置于
ComfyUI/models/text_encoders/umt5_xxl_fp8_e4m3fn_scaled.safetensors
FP8 缩放文本编码器。放置于
ComfyUI/models/text_encoders/wan_2.1_vae.safetensors
Wan2.1 VAE 模型。放置于
ComfyUI/models/vae/clip_vision_h.safetensors
用于图像条件的 CLIP Vision 模型。放置于
ComfyUI/models/clip_vision/对于 diffusion 模型,本指南将使用 fp16 精度模型,因为我们发现它们比 bf16 版本表现更好。如果你需要其他精度的版本,请访问这里下载。
Wan2.1 Text-to-Video Workflow (1.3B)
Wan 2.1 Text to Video
Generate videos from text prompts using Wan 2.1.
Run on Comfy Cloud
Open in Comfy Cloud
Download Workflow
Download JSON or search “Wan 2.1 Text to Video” in Template Library
Model Downloads
wan2.1_t2v_1.3B_fp16.safetensors
Diffusion model for Wan2.1 Text-to-Video. Place in
ComfyUI/models/diffusion_models/If you need other t2v precision versions, please visit here to download them.
Steps to Run

- Make sure the
Load Diffusion Modelnode has loaded thewan2.1_t2v_1.3B_fp16.safetensorsmodel - Make sure the
Load CLIPnode has loaded theumt5_xxl_fp8_e4m3fn_scaled.safetensorsmodel - Make sure the
Load VAEnode has loaded thewan_2.1_vae.safetensorsmodel - (Optional) You can modify the video dimensions in the
EmptyHunyuanLatentVideonode if needed - (Optional) If you need to modify the prompts (positive and negative), make changes in the
CLIP Text Encodernode at number5 - Click the
Runbutton or use the shortcutCtrl(cmd) + Enterto execute the video generation
Wan2.1 Image-to-Video Workflow (14B)
Since Wan Video separates the 480P and 720P models, we’ll need to provide examples for both resolutions in this guide. In addition to using different models, they also have slight parameter differences.480P Version
Wan 2.1 Image to Video
Generate videos from images using Wan 2.1.
Run on Comfy Cloud
Open in Comfy Cloud
Download Workflow
Download JSON or search “Wan 2.1 Image to Video” in Template Library
LoadImage node:
image_to_video_wan_start_image.png
LoadImage node 52 · image_to_video_wan_start_image.png
Model Downloads
wan2.1_i2v_480p_14B_fp16.safetensors
Diffusion model for Wan2.1 I2V 480P. Place in
ComfyUI/models/diffusion_models/Steps to Run

- Make sure the
Load Diffusion Modelnode has loaded thewan2.1_i2v_480p_14B_fp16.safetensorsmodel - Make sure the
Load CLIPnode has loaded theumt5_xxl_fp8_e4m3fn_scaled.safetensorsmodel - Make sure the
Load VAEnode has loaded thewan_2.1_vae.safetensorsmodel - Make sure the
Load CLIP Visionnode has loaded theclip_vision_h.safetensorsmodel - Upload the provided input image in the
Load Imagenode - (Optional) Enter the video description content you want to generate in the
CLIP Text Encodernode - (Optional) You can modify the video dimensions in the
WanImageToVideonode if needed - Click the
Runbutton or use the shortcutCtrl(cmd) + Enterto execute the video generation
720P Version
Run on Comfy Cloud
Open in Comfy Cloud
Download Workflow
Download the workflow image and drag it into ComfyUI to load the workflow
Input Image
Download the default input image, or use your own image.
Model Downloads
wan2.1_i2v_720p_14B_fp16.safetensors
Diffusion model for Wan2.1 I2V 720P. Place in
ComfyUI/models/diffusion_models/Steps to Run

- Make sure the
Load Diffusion Modelnode has loaded thewan2.1_i2v_720p_14B_fp16.safetensorsmodel - Make sure the
Load CLIPnode has loaded theumt5_xxl_fp8_e4m3fn_scaled.safetensorsmodel - Make sure the
Load VAEnode has loaded thewan_2.1_vae.safetensorsmodel - Make sure the
Load CLIP Visionnode has loaded theclip_vision_h.safetensorsmodel - Upload the provided input image in the
Load Imagenode - (Optional) Enter the video description content you want to generate in the
CLIP Text Encodernode - (Optional) You can modify the video dimensions in the
WanImageToVideonode if needed - Click the
Runbutton or use the shortcutCtrl(cmd) + Enterto execute the video generation