Inputs
Note: If you provide a
start_image, you must also connect a vae so it can be encoded. The start_image is automatically upscaled to 16 times the spatial dimensions (width and height) of the input latent, then encoded and placed into the conditioning latent. Only the RGB channels of the start_image are used for encoding.
Outputs
This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub
Source fingerprint (SHA-256):
c9e64092e78423f5e0dc43446a77240e09100242c25e4fccc91491049fe76be5