Skip to main content
This node creates an empty latent that combines both video and audio information for the MiniMax H3 model. You define the width, height, and length of the content, and the node produces a blank latent that the model can use as a starting point for generation. The duration (length) is automatically adjusted to fit the model’s required frame grid of 17k+5 frames at 24 fps.

Inputs

Note: The length value is snapped up to the next frame count that fits the model’s 17k+5 grid (17 x k + 5 frames, such as 5, 22, 39, 56, 73, 90, 107, 124, and so on). The width and height values must be multiples of 32. The maximum resolution is the system-defined value in ComfyUI.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): ee24f4ac630858d87b9b98bb402689a5790e0ed882ec47dffe7b497216e37a5c