Is FP8 Quant a Bad Idea for Qwen 3.8 27B? Quality Concerns Raised
lblblllb · reddit · 2026-08-18
A user questions the quality of the official FP8 quantization for Qwen 3.8 27B, noting that while it matches the size of Q80, it shows significantly worse KL divergence in tests. The discussion asks if others have switched away from FP8 for better performance.
Related event: Qwen3.8-27B FP8 quantization quality questioned by users(2 posts)→
More from Infra
- Rust Inference Engine Core: 3k Lines, Minimal Tech Debt, Fast Startup — charles_irl · 2026-08-18
- Fal seeks 128 nodes of AMD MI350/355 compute — isidentical · 2026-08-18
- US labs spend 10x more on compute than China, but China builds faster and has easier energy access — Jsevillamol · 2026-08-18
- Tidebroker: Secure Credential Brokerage for AI Agents — steipete · 2026-08-18
- AI Token usage up 1139%, costs down 55% amid caching shift — AccBalanced · 2026-08-18
- US power shortage to leave 43 GW of AI chips idle by 2030 — AccBalanced · 2026-08-18