



Qwen-Image-2.1 is open-weight. Download it from Hugging Face (Qwen/Qwen-Image-2.1) or ModelScope, or pull the ComfyUI-ready copy from Comfy-Org/Qwen-Image-2.1.
Diffusers supports it from day one through QwenImage21Pipeline, and ComfyUI ships native text-to-image and editing workflows. For heavier use, vLLM-Omni, SGLang, and LightX2V add prefix caching, quantisation, and multi-GPU inference; AMD ROCm and FlagOS chips are supported too.
Start from the official prompt-rewriting checkpoints (PE-T2I for generation, PE-I2I for editing), keep to the recommended 2K sizes and 40 steps, and for a transparent result open the prompt with the RGBA phrasing the model expects.
Qwen-Image-2.1 is open-weight and runs on your own hardware. If you would rather write a prompt and get an image back, open Qwen Image 3 on GoEnhance and start from one of the examples above.
Start with Qwen Image 3