Skip to main content
The CLIPTextEncodeHiDream node processes four separate text inputs using different language models (CLIP-L, CLIP-G, T5-XXL, and LLaMA) and combines them into a single conditioning output. It tokenizes each text input with its corresponding model and encodes them together using a scheduled encoding approach, enabling more sophisticated text conditioning by leveraging multiple language models simultaneously.

Inputs

Note: All four text inputs (clip_l, clip_g, t5xxl, and llama) are required for proper functioning, as each contributes to the final conditioning output through the scheduled encoding process.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): c5e269c17bd2dd7d7171c02598a87983a988d953dd7df285978fc25a9c896e46