Qwen Image 2.1 quantized to 2x speed — but INT4 quantization loses to FP4

mesmerlord · reddit · 2026-09-22

A developer quantized the new Qwen Image 2.1 to replicate their paid tool's serverless setup (previously Flux klein + nunchaku), publishing weights as Mesmer-Image-21-Nunchaku on Hugging Face.

Key findings:

Original post →

More from Infra

Infra channel →