Skip to main content
This node prepares the conditioning and empty latent needed to generate a video with the MiniMax H3 model. It takes a text prompt and, optionally, images for the first and/or last frame of the video, and converts them into model inputs. Keyframe images are resized, encoded, and attached to the conditioning at the start and end of the video.

Inputs

When first_frame and/or last_frame are provided, the keyframe images are encoded with the VAE and attached to the conditioning at frame 0 and at the final frame, respectively. When neither is provided, the node works from the prompt alone.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 46efc87bd46f4a86cb6df37c75f960419a2a98b34480e7dc0023c9d87903870b