Alibaba Qwen Releases Qwen-Image-2.1: A 7B Open-Weight Model for Image Generation and Editing
Alibaba’s Qwen team released Qwen-Image-2.1, a 7B open-weight model for image generation and editing.
Alibaba’s Qwen team released Qwen-Image-2.1, a unified text-to-image and editing model whose diffusion transformer has 7 billion parameters across 32 single-stream DiT layers, down from the 20B Qwen-Image shipped in August 2025. The pipeline also loads an 8B Qwen3-VL encoder and a 64-channel RGBA VAE; one checkpoint handles generation, multi-reference editing of up to 10 images, local edits, and native transparency, defaulting to 2048×2048. On Qwen’s in-house Qwen-Image-Bench it scores 60.28, above listed open-weight models including FLUX 2 Max at 55.33, but below GPT Image 2.5 Sunburst at 67.01. Day-0 tooling includes Diffusers, ComfyUI, vLLM-Omni, and SGLang; commercial use requires a separate Qwen license.