the AI bench
VERIFIED SEPTEMBER 2026
All fast takes

FAST TAKE · 2026-09-20 · QWEN-IMAGE-2.1

Qwen-Image-2.1 shipped: unified generation and editing, native RGBA — and a licence trap

Alibaba released Qwen-Image-2.1 on September 20, 2026. It is the runnable successor to the announced-but-never-released 2.0: a 7B visual generator with one pipeline for text-to-image and editing, native 2K output, transparent RGBA generation, and support for up to ten reference images. The important correction is legal, not technical: the repository uses the Qwen Research License, not Apache 2.0, and commercial use requires a separate licence.

Verdict: the promised unified Qwen image model finally ships — excellent text, editing and native transparency, but the weights are research-only


The take

From Qwen's release and model card: a 32-block single-stream diffusion transformer, mixed-granularity attention and prefix-KV reuse, generation and instruction editing in the same model, precise paint-and-mask edits, identity preservation, multilingual typography, native transparent backgrounds and up to ten reference images. The official examples run at native 2K dimensions. The Hugging Face repository had 6,523 downloads and 1,523 likes when we checked on September 22; a ComfyUI-packaged community repository had already passed 500K downloads.

The hardware story needs restraint. The visual transformer itself is roughly 14.2 GB in the official repository, but the full pipeline also includes the Qwen3-VL text encoder and VAE; the complete download is about 39 GB. Qwen documents CPU offload, not an official VRAM table. Community quantized and ComfyUI paths exist, but we will not turn one anecdotal memory figure into a buying recommendation. Treat 24 GB as the comfortable local target and expect component offload below that.

The licence is decisive. The actual LICENSE file permits non-commercial research and evaluation; commercial use requires a separate agreement from Qwen. That means 2.1 becomes our strongest research/non-commercial Qwen image pick, while Qwen-Image-2512 remains the commercially clean Apache-2.0 choice. We updated the model page, image planner and use-case copy to keep both truths visible.

Where this fits

Models: Qwen-Image-2.1 (7B) + Qwen-Image-2512 (20B) · HiDream-O1-Image (8B) · FLUX.2 [dev] · Z-Image-Turbo

Hardware: NVIDIA RTX 5090 · NVIDIA RTX 4090 · M5 Max MacBook Pro 64 GB

Sources

Next step

Try this in the planner→