Skip to main content
Text Encode Ming Image Edit encodes a text prompt into conditioning for Ming image editing, optionally mixing in reference images. The prompt and the reference images are tokenized by a CLIP model, and when a VAE is connected the reference images are also encoded into latent frames that are appended to the conditioning.

Inputs

Note: Reference images are read in numeric order of their slot names, and empty slots are ignored. Reference latents are only produced when both vae and at least one image are provided.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 675fb3cc0af006e1284fdb2a5ca2c268c540ee92e1ede90b2359f4e3fc8ba5ea