Ternary Bonsai 2 27B: 5.9GB weights retain ~95% of full-precision reasoning

cephaloform · x · 2026-09-23

Ternary Bonsai 2 27B is out: ternary-weight compression shrinks a 27B model to just 5.9GB of weights, versus 54GB for full-precision Qwen3 27B.

The author benchmarked how much reasoning survived compression against full-precision Qwen3 27B and Gemma 4 12B QAT (7GB):

The writeup discusses where compression preserves performance and where gaps remain.

Original post →

More from Infra

Infra channel →