Unsloth Releases Qwen2.5-72B Quants: 1-bit Version Runs on 8GB RAM

cephaloform · x · 2026-08-20

Unsloth AI has released new GGUF quantizations for Qwen2.5-72B. The Dynamic V3 version achieves >10% higher accuracy on benchmarks like Div-300 and KLD. Notably, they also released a 1-bit extreme quantization that runs on just 8GB RAM while retaining approximately 77% of BF16 accuracy.

Original post →

More from Infra

Infra channel →