Qwen 27B NVFP4 Quantization Underperforms Baseline
Pyrolistical · reddit · 2026-09-01
The author evaluated Qwen3.8 27B NVFP4 GGUF models against Unsloth baselines using perplexity scores. Results show that all tested NVFP4 models have worse perplexity for their file size compared to standard quants, even underperforming Q4KXL. This likely explains why Unsloth does not publish NVFP4 GGUF versions.
More from Models
- Testing unreleased Gemini 3.8 Flash: no citations shown for top-of-funnel queries — gaganghotra_ · 2026-09-03
- Hidden-bug eval across 105 issues: Fable 5.1 finds 43, none fixes all — cost per model compared — PawelHuryn · 2026-09-03
- X open-sources new For You algorithm code: long dwell drives retrieval, bots can trigger account review — Kyrannio · 2026-09-03
- Gemini 3.8 Flash reverse-engineers Kerbal save files to build and land a Mun rocket — dosco · 2026-09-03
- User calls out model for double-standard answers on gendered scenario questions — Ribbitz_bow_tie27 · 2026-09-03
- Redditor Predicts Astra Model Release Tomorrow at 1pm PT Based on X Teasers — dolo937 · 2026-09-03