Qwenโs Image 2.1 released but it come with lots of improvements but just as many caveats.
The old Qwen-Image-2512 was a 20B model with a massive:
๐พ 40.9GB BF16 image transformer
But the new Qwen-Image-2.1 comes with โฆ
๐ง 7B visual generator
๐พ 14.2GB BF16
๐ฅ 7.26GB INT8 already available for ComfyUI ๐๐ (Day-1)
This little thing can โฆ
๐จ generate AND edit images
๐ผ๏ธ generate native 2K
๐ซฅ create real transparent RGBA images
๐ฅ use up to 10 reference images
โ๏ธ preserve people/products while editing
๐ค render text
๐ฏ do masked/local edits
Qwen basically took the huge local image model and reduced with a catch โฆ
โ ๏ธ The tiny 7.26GB image model comes with an entire pipeline.
It uses a Qwen3-VL 8B encoder + VAE.
ComfyUI already has a 6.31GB W4A8 encoder, and Qwen supports CPU offloading for smaller GPUs.
Alsoโฆ
๐ซ Qwen-Image-2.1 is NON-COMMERCIAL under the new Qwen Research License.
Thatโs especially notable because Qwen-Image-2512 was Apache 2.0.