Ternary Bonsai 2 27B released: 9x smaller, 98.2% of full precision, just 5.9 GB under Apache 2.0

cephaloform · x · 2026-09-18

PrismML announced Ternary Bonsai 2 27B, a ternary-quantized model based on Qwen3.8 27B that is 9x smaller than its full-precision counterpart while retaining 98.2% of aggregate benchmark performance. Two months after the first Bonsai 27B, the 5.9 GB footprint is unchanged but quality gaps narrowed sharply, with strong gains in agentic coding, multimodal reasoning, and long-horizon tool use. Released under Apache 2.0; users are already eyeing running it locally on a 3080 and canceling API subscriptions.

Related event: PrismML's Ternary Bonsai 2 27B Shrinks to 5.9GB, Retains 98.2% Performance(20 posts)→

Original post →

More from Infra

Infra channel →