Skip to main content
The HunyuanVideo15SuperResolution node prepares conditioning data for a video super-resolution process. It takes a latent representation of a video and, optionally, a starting image, and packages them along with noise augmentation and CLIP vision data into a format that can be used by a model to generate a higher-resolution output.

Inputs

Note: If you provide a start_image, you must also connect a vae so it can be encoded. The start_image is automatically upscaled to 16 times the spatial dimensions (width and height) of the input latent, then encoded and placed into the conditioning latent. Only the RGB channels of the start_image are used for encoding.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): c9e64092e78423f5e0dc43446a77240e09100242c25e4fccc91491049fe76be5