Skip to main content
Generate text responses with Google’s Gemini models. Provide a text prompt and, optionally, one or more images, audio clips, videos, or files as multimodal context. The node sends the prompt and any attached media to the selected model and returns the model’s text answer. Note: This node is marked as deprecated in the source code.

Inputs

Common Inputs

Gemini 3.8 Flash Inputs

These inputs appear when model is set to "Gemini 3.8 Flash". Note: This model does not expose temperature or top_p sampling controls.

Gemini 3.7 Flash Inputs

These inputs appear when model is set to "Gemini 3.7 Flash".

Gemini 3.5 Flash Inputs

These inputs appear when model is set to "Gemini 3.5 Flash".

Gemini 3.1 Pro Inputs

These inputs appear when model is set to "Gemini 3.1 Pro".

Gemini 3.1 Flash-Lite Inputs

These inputs appear when model is set to "Gemini 3.1 Flash-Lite".

Media and File Inputs

The following inputs are shared by all models and appear alongside the model-specific inputs. Note: When media (images, audio, or video) is attached, the node uploads the first 10 media items to ComfyAPI storage and passes them as URLs; this URL budget is shared across all media types and is consumed in order (video first, then audio, then images). Any remaining media is encoded inline as base64 data, with a maximum combined inline payload of 18 MB. If the inline payload would exceed 18 MB, the node raises an error. The prompt parameter must contain at least one non-whitespace character. Setting seed to 0 requests a random seed.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 8ae14c6465569695e1e99b0040cb2c745e5f3d9ddcd3b03013530dc7e7135b87