Two RTX 3090s still struggle to fit Qwen Image Edit alongside a 27B text model

Civil_Fee_7862 · reddit · 2026-07-30

Running Qwen Image Edit on two 3090s still hits VRAM limits

A user describes attempts to support image-generation tasks on a dual-RTX-3090 setup while also running a 4-bit Qwen3.6-27B text model with an 8-bit KV cache.

What they tried

Where they ended up

The author is considering buying a third RTX 3090 dedicated to image generation, because even if a 4-bit Qwen Image Edit model can be squeezed into roughly 14 GB, an 8-bit version likely will not fit. They ask whether anyone has built a similar multi-GPU setup, and whether there are PyTorch settings that make tensor or pipeline parallelism behave reliably for stable-diffusion-like models such as Qwen Image Edit.

Original post →

More from Infra

Infra channel →