Skip to main content
The WanPhantomSubjectToVideo node prepares conditioning data and a latent for Wan video generation. It creates an empty latent video from the requested width, height, length, and batch size, and, when reference images are supplied, encodes them with the VAE and adds them to the conditionings as time-dimensional visual guidance.

Inputs

Note: When images are provided, they are automatically upscaled to match the specified width and height, and only the first length images are used for processing. Each image is encoded with the vae and concatenated along the time dimension, and only the RGB channels of each image are used.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): a1853382f6e564f66262b69dd7b06cc58e26b93386a460a98e6fcc2ff6acf12b