Calibrated Qwen3.6-27B quantization tests weight groups before compressing them

enginetown · reddit · 2026-07-28

A solo builder released Qwen3.6-27B-Calibrated after testing which weight groups can be compressed safely before quantization. The project measures divergence one weight group at a time with KL divergence across general, code, math, and tool-use prompts, instead of applying a flat bit-depth everywhere.

The author says the results show a narrow quality cliff around 3.5–3.9 bits per weight group, and that tool calling tends to break first. Three builds were published: Bedrock at 13.26 GB, Tightrope at 12.53 GB, and Gambit at 10.94 GB.

Original post →

More from Infra

Infra channel →