SOTA GGUF quants for Qwen3.8-Flash-Next: near-baseline quality at 1/5 the size

BullfrogScary8947 · reddit · 2026-09-16

ISTA-DASLab released GSQ-RCO GGUF quantizations of Qwen3.8-Flash-Next, cutting the 80-95GB model to 68-76GB with near-baseline quality. Key points:

Available on Hugging Face for local deployment, pick per throughput/quality needs.

Original post →

More from Infra

Infra channel →