Alleged deceptive quantization: AtomicChat accused of faking Q4 levels

po_stulate · reddit · 2026-09-01

A user alleges that AtomicChat's Qwen3.8-Flash-Next Q4KM quant is deceptive, noting the file size is suspiciously small (56GB). Inspection reveals most tensors are IQ2S instead of Q4K, and the GGUF metadata indicates IQ2S. Despite a high KLD score on the model card, the author suspects a low-precision model is being misrepresented as a higher-precision quant.

Original post →

More from Models

Models channel →