Skip to main content
The Wan22FunControlToVideo node prepares conditioning data and an empty latent tensor for video generation with the Wan video model. It encodes optional reference images and control videos into latent space, attaches them to the positive and negative conditioning, and creates a zero-filled latent tensor with the correct spatial and temporal dimensions for the requested video.

Inputs

Note: The length parameter is processed in steps of 4 frames, and the node automatically applies temporal scaling when building the latent space. When ref_image is provided, only its first frame is encoded and attached to the conditioning as reference latents. When control_video is provided, it is trimmed to length frames, encoded, and placed into the concat latent used by the conditioning. The start_image parameter is referenced in the execution logic but is not exposed in the node’s input schema.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 731b848f15c13ddc662f19230acb55d195f934bad7d9ae516a288e0ed8f8d899