Skip to main content
The TextEncodeQwenImage21 node encodes a prompt and a negative prompt for the Qwen-Image 2.1 model, optionally attaching reference images. Reference images are seen by the text encoder and, when a VAE is connected, are also encoded as latents that are spliced into the sequence, so the conditioning carries both the text instruction and the visual reference. The node returns positive and negative conditioning together with an empty latent sized to the first reference image, ready to be sampled.

Inputs

The empty latent output is sized to the first connected reference image, or to resolution when no reference image is connected. Sample at the latent this node returns: any other size shifts the edit. When a VAE is connected, the same reference latents are attached to both the positive and the negative conditioning, so a single sampler step can denoise both.

Outputs

This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub

Source fingerprint (SHA-256): 3870f04597d12b593498c12ca139428af2d717b65aa40c889de9373ae9eb475e