Qwen3.8 Flash AP quants beat other high-quality quants with new KLD eval method
Dutchnamn · reddit · 2026-09-03
The agentionai team released AP (high-precision) GGUF quants of Qwen3.8-Flash-Next, and benchmarking shows they beat other high-quality quants.
- They had to devise a modified KLD measurement with a new dataset, since the NGRAM dataset effectively memorized all of Wikipedia and skewed evaluation
- The quants balance high precision with prefill performance
- Full model card is on Hugging Face; community feedback welcome
More from Models
- AI vividly 'sees' and describes scenes while experiencing only darkness — yeastsplainer · 2026-09-03
- Every's writing bench adds Gemini 3.8 Flash, Grok 4.6, and Muse Spark 1.3 — danshipper · 2026-09-03
- Meta's Muse Spark 1.3 lands on OpenRouter with 1M context for agentic workflows — armand_ruiz · 2026-09-03
- DeepSeek-V4-Pro ships with 1.6T-param MoE; open-source eval harness steals the show — DeepLearningAI · 2026-09-03
- Rival AI agents: cross-vendor model review catches what self-review misses — rseroter · 2026-09-03
- Gemini's Distinctive Take on AI Sentience Turns Heads — aiamblichus · 2026-09-03