Skip to main content
Trellis2Conditioning converts an input image into conditioning data for the TRELLIS.2 model. It uses a CLIP vision model to encode the image into two sets of features (at 512 and 1024 scales) and packages them as a positive conditioning pair, while also creating a matching zero-filled negative conditioning pair that serves as an empty reference.

Inputs

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 467698e58558ceca9ac633d63aacf360a1eb674ac4ebd47de7423f85e62c0fe6