Community quants for Qwen3.8 Flash save 20-30GB at same quality as unsloth

Dutchnamn · reddit · 2026-08-29

A Reddit user released a set of Qwen3.8 Flash (Next) GGUF quants after days of benchmarking. They require 20-30GB less disk and RAM than comparable unsloth or AesSedai quants at equal quality. PPL is published in the readme and is competitive; both Q4 quants are strong, with Q3/Q5 to follow. Built with imatrix and a tailored quantization recipe. A ROCmFP4 variant for AMD users is slightly better and faster than Q4XS.

Original post →

More from Infra

Infra channel →