Inputs
Reference latents are added to both conditioning outputs only when
vae is provided together with at least one reference image. If vae is omitted, the positive output still receives vision tokens from the reference images, but neither output includes reference latents.
Outputs
This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub
Source fingerprint (SHA-256):
170979acf5b2e9f25f96231a4b23a4376cfddcd4bda2fdd6e03528417e6931b0