Skip to main content
Wan2.1 Video series is a video generation model open-sourced by Alibaba in February 2025 under the Apache 2.0 license. It offers two versions:
  • 14B (14 billion parameters)
  • 1.3B (1.3 billion parameters) Covering multiple tasks including text-to-video (T2V) and image-to-video (I2V). The model not only outperforms existing open-source models in performance but more importantly, its lightweight version requires only 8GB of VRAM to run, significantly lowering the barrier to entry.
Make sure your ComfyUI is updated.Workflows in this guide can be found in the Workflow Templates. If you can’t find them in the template, your ComfyUI may be outdated.If nodes are missing when loading a workflow, possible reasons:
  1. You are not using the latest ComfyUI version (Nightly version)
  2. Some nodes failed to import at startup

Wan2.1 ComfyUI Native Workflow Examples

Please update ComfyUI to the latest version before starting the examples to make sure you have native Wan Video support.

Model Installation

All models mentioned in this guide can be found here. Below are the common models you’ll need for the examples in this guide, which you can download in advance: Choose one version from Text encoders to download:

umt5_xxl_fp16.safetensors

FP16 precision text encoder. Place in ComfyUI/models/text_encoders/

umt5_xxl_fp8_e4m3fn_scaled.safetensors

FP8 scaled text encoder. Place in ComfyUI/models/text_encoders/
VAE

wan_2.1_vae.safetensors

Wan2.1 VAE model. Place in ComfyUI/models/vae/
CLIP Vision

clip_vision_h.safetensors

CLIP Vision model for image conditioning. Place in ComfyUI/models/clip_vision/
File storage locations:
For diffusion models, we’ll use the fp16 precision models in this guide because we’ve found that they perform better than the bf16 versions. If you need other precision versions, please visit here to download them.

Wan2.1 Text-to-Video Workflow (1.3B)

Wan 2.1 Text to Video

Generate videos from text prompts using Wan 2.1. Wan 2.1 Text to Video workflow preview

Run on Comfy Cloud

Open in Comfy Cloud

Download Workflow

Download JSON or search “Wan 2.1 Text to Video” in Template Library

Model Downloads

wan2.1_t2v_1.3B_fp16.safetensors

Diffusion model for Wan2.1 Text-to-Video. Place in ComfyUI/models/diffusion_models/
If you need other t2v precision versions, please visit here to download them.

Steps to Run

ComfyUI Wan2.1 Workflow Steps
  1. Make sure the Load Diffusion Model node has loaded the wan2.1_t2v_1.3B_fp16.safetensors model
  2. Make sure the Load CLIP node has loaded the umt5_xxl_fp8_e4m3fn_scaled.safetensors model
  3. Make sure the Load VAE node has loaded the wan_2.1_vae.safetensors model
  4. (Optional) You can modify the video dimensions in the EmptyHunyuanLatentVideo node if needed
  5. (Optional) If you need to modify the prompts (positive and negative), make changes in the CLIP Text Encoder node at number 5
  6. Click the Run button or use the shortcut Ctrl(cmd) + Enter to execute the video generation

Wan2.1 Image-to-Video Workflow (14B)

Since Wan Video separates the 480P and 720P models, we’ll need to provide examples for both resolutions in this guide. In addition to using different models, they also have slight parameter differences.

480P Version

Wan 2.1 Image to Video

Generate videos from images using Wan 2.1. Wan 2.1 Image to Video workflow preview

Run on Comfy Cloud

Open in Comfy Cloud

Download Workflow

Download JSON or search “Wan 2.1 Image to Video” in Template Library
Input materials Upload this file to the matching LoadImage node:

image_to_video_wan_start_image.png

LoadImage node 52 · image_to_video_wan_start_image.png
image_to_video_wan_start_image.png

Model Downloads

wan2.1_i2v_480p_14B_fp16.safetensors

Diffusion model for Wan2.1 I2V 480P. Place in ComfyUI/models/diffusion_models/

Steps to Run

ComfyUI Wan2.1 Workflow Steps
  1. Make sure the Load Diffusion Model node has loaded the wan2.1_i2v_480p_14B_fp16.safetensors model
  2. Make sure the Load CLIP node has loaded the umt5_xxl_fp8_e4m3fn_scaled.safetensors model
  3. Make sure the Load VAE node has loaded the wan_2.1_vae.safetensors model
  4. Make sure the Load CLIP Vision node has loaded the clip_vision_h.safetensors model
  5. Upload the provided input image in the Load Image node
  6. (Optional) Enter the video description content you want to generate in the CLIP Text Encoder node
  7. (Optional) You can modify the video dimensions in the WanImageToVideo node if needed
  8. Click the Run button or use the shortcut Ctrl(cmd) + Enter to execute the video generation

720P Version

Wan2.1 Image-to-Video Workflow 14B 720P

Run on Comfy Cloud

Open in Comfy Cloud

Download Workflow

Download the workflow image and drag it into ComfyUI to load the workflow

Input Image

Download the default input image, or use your own image.

Model Downloads

wan2.1_i2v_720p_14B_fp16.safetensors

Diffusion model for Wan2.1 I2V 720P. Place in ComfyUI/models/diffusion_models/

Steps to Run

ComfyUI Wan2.1 Workflow Steps
  1. Make sure the Load Diffusion Model node has loaded the wan2.1_i2v_720p_14B_fp16.safetensors model
  2. Make sure the Load CLIP node has loaded the umt5_xxl_fp8_e4m3fn_scaled.safetensors model
  3. Make sure the Load VAE node has loaded the wan_2.1_vae.safetensors model
  4. Make sure the Load CLIP Vision node has loaded the clip_vision_h.safetensors model
  5. Upload the provided input image in the Load Image node
  6. (Optional) Enter the video description content you want to generate in the CLIP Text Encoder node
  7. (Optional) You can modify the video dimensions in the WanImageToVideo node if needed
  8. Click the Run button or use the shortcut Ctrl(cmd) + Enter to execute the video generation