Skip to main content
CosmosPredict2ImageToVideoLatent creates video latent representations from images for video generation. It can generate a blank video latent or incorporate start and end images to create video sequences with specified dimensions and duration. The node handles the encoding of images into the appropriate latent space format for video processing.

Inputs

Note: When neither start_image nor end_image are provided, the node generates a blank video latent. When one or both images are provided, they are resized to width and height, encoded into latent space, and positioned at the beginning and/or end of the video sequence, with the corresponding regions marked in the noise mask so they are preserved during generation. The resulting latent and mask are repeated batch_size times.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 842bd2b8cda438e7b938439d4eba280478939e3302dc1846d52595d40082ff05