Calibration-Free Quantization Method TQ Open-Sourced, Hits 92.4% Top-1 on Qwen 27B 4-bit
textclf · reddit · 2026-09-25
TextCLF released TQ, a calibration-free quantization method that needs no data, so models can be quantized the day they launch. Its 4-bit Qwen 3.8 27B shows mean KLD of 0.0282 and 92.4% top-1, close to calibration-based methods. The code is open-sourced as Quant Factory (GitHub), with quantized models on Hugging Face and a Docker image for vLLM serving. Currently 4-bit only, with 2-bit and 3-bit planned.
More from Infra
- AI data center talent war: electrician pay up 3-5x in Texas and Louisiana, says SemiAnalysis — johncoogan · 2026-09-25
- CoreWeave turns away $100M customers until May as Nebius climbs to Platinum GPU cloud tier — johncoogan · 2026-09-25
- Inside Quail: custom vLLM scheduler, workload-aware KV cache for 1B tok/min — sh_reya · 2026-09-25
- Anthropic's CI job volume grew 25x in six months — here's how they scaled test selection — JeremyCMorgan · 2026-09-25
- Chip startups like Cerebras and Groq are becoming the next wave of neoclouds, says SemiAnalysis — johncoogan · 2026-09-25
- Anthropic Commits $11.6B Over 7 Years to Akamai Cloud, Validating Distributed Inference Thesis — pdamodaran · 2026-09-25