Calibrated Qwen3.6-27B quantization tests weight groups before compressing them
enginetown · reddit · 2026-07-28
A solo builder released Qwen3.6-27B-Calibrated after testing which weight groups can be compressed safely before quantization. The project measures divergence one weight group at a time with KL divergence across general, code, math, and tool-use prompts, instead of applying a flat bit-depth everywhere.
The author says the results show a narrow quality cliff around 3.5–3.9 bits per weight group, and that tool calling tends to break first. Three builds were published: Bedrock at 13.26 GB, Tightrope at 12.53 GB, and Gambit at 10.94 GB.
More from Infra
- AI is likely to control quantum computers first, then use them for narrow science tasks — imjustnewatai · 2026-07-28
- Moonshot’s Kimi K3 lands in Japan with 2.8T open weights and $3/$13 pricing — DavidBennett__ · 2026-07-28
- A Kimi-k3 joke contrasts a $1,908 annual plan with $1.0908M to run it at home — HarveenChadha · 2026-07-28
- Frozen 12B system reuses verified memory at zero tokens and 6,000,000-token context — Corbenci · 2026-07-28
- OpenAI’s expected $750 billion compute spend puts Anthropic’s leasing strategy under pressure — remybigot · 2026-07-28
- AMD, SGLang and Moonshot ship together as the chip wars shift to infrastructure — AnushElangovan · 2026-07-28