Skip to main content
Generate a video with audio from a text prompt using Google’s Gemini Omni Flash model. Optionally provide reference images and/or videos to guide or edit the result. Describe the desired length (3-10s) and aspect ratio (16:9 or 9:16) directly in the prompt.

Inputs

Notes:
  • If an image input contains multiple frames, each frame counts toward the maximum of 14 images.
  • When images or videos are provided, the combined encoded media size must stay under about 90 MB; otherwise the node raises an error.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 1b7ca51d07cfb6a166cfed2a7e7174fd62f3290abcc1bdfdce94369dda242d3f