GLM-5.2 Weights Reduced by 423GB

EAccelerate_42 · x · 2026-07-13

The author claims to have reduced the size of **GLM-5.2** by **423GB** **without altering the model**: - Reduced from **1403GB** to **980GB** - Model size is **753B weights** - Results remain **bit-for-bit exact** - **No quantization and no retraining** The key insight is that the weights remain in a compressed state within **VRAM**, rather than reconstructing the full model first. The author mentions the full write-up and repo will be released in the next post.

Original post →

More from Infra

Infra channel →