Qwen-Image 2.1: open-weights 7B unified gen+edit model, 30s on a 5080 laptop
Lexius2129 · reddit · 2026-09-21
Qwen-Image 2.1 shipped with Day-0 ComfyUI support. Built on February's Qwen-Image 2.0 architecture but released as open weights (research-only license):
- Unified generation and editing in one compact 7B model; up to 10 input images
- Much better textures than prior Qwen-Image-Edit; strong text rendering for infographics, posters, Chinese calligraphy
- Native 2K output and alpha-channel/transparency generation
Benchmarks on a 5080 laptop (16GB VRAM, 32GB RAM) with qwenimage2.1int8convrot: 32s average with the int8 text encoder, 36s with Kijai's w4a8 quantized encoder, 46s with bf16 — and 30s with w4a8 plus the 'Model Attention Backend' node set to comfy kitchen attention.
To try it: update ComfyUI or use Comfy Cloud, grab weights from Comfy-Org/Qwen-Image-2.1 on Hugging Face, and load the text-to-image or image-edit workflow templates.
More from Multimodal
- Creator explores color cycling in AI video with a YouTube long-form scene-to-palette glide — PurzBeats · 2026-09-21
- AI-Generated Retro Anime OVA Squad37 Returns With Episode 5 — Slow_Opposite · 2026-09-21
- Reddit Users Say Qwen 2512 Still King of Photorealism Despite Flaws — Inevitable_Pen9043 · 2026-09-21
- Turning GPT-6 Astra Into a Motion Graphics Studio: 5 Ad Renders in 18 Minutes — PrajwalTomar_ · 2026-09-21
- Running MiniMax H3 on a 4060 Ti 16GB: 500s for a 5-second clip, author seeks speedups — GranDaddyP · 2026-09-21
- H3 v2v motion reference beats prompt-only i2v for dance MV, creator says — R34vspec · 2026-09-21