Run Qwen3.8 Locally: Unsloth Shrinks Model Size by 91% via 1-bit Quantization

danielhanchen · x · 2026-08-13

Unsloth AI announced that they have successfully shrunk the massive Qwen3.8-2.4T-A95B model from 4.9TB to 397GB (a 91% reduction) using their new Dynamic 1-bit quantization, enabling local execution.

Related event: Unsloth Enables Local Qwen3.8 Deployment via 1-bit Quantization(2 posts)→

Original post →

More from Infra

Infra channel →