Calibration-Free Quantization Method TQ Open-Sourced, Hits 92.4% Top-1 on Qwen 27B 4-bit

textclf · reddit · 2026-09-25

TextCLF released TQ, a calibration-free quantization method that needs no data, so models can be quantized the day they launch. Its 4-bit Qwen 3.8 27B shows mean KLD of 0.0282 and 92.4% top-1, close to calibration-based methods. The code is open-sourced as Quant Factory (GitHub), with quantized models on Hugging Face and a Docker image for vLLM serving. Currently 4-bit only, with 2-bit and 3-bit planned.

Original post →

More from Infra

Infra channel →