PrismML's Ternary Bonsai 2 27B is 9x smaller at 5.9GB while keeping 98.2% benchmark performance

alexcovo_eth · x · 2026-09-19

PrismML released Ternary Bonsai 2 27B, a ternary-quantized version of Qwen3.8 27B that is 9x smaller at 5.9GB while retaining 98.2% of the full-precision model's aggregate benchmark performance.

The biggest change over the first Bonsai 27B released two months ago is quality: same footprint, but a materially narrower gap to full precision, with strong gains in agentic coding, multimodal reasoning, and long-horizon tool use. Available today under Apache 2.0 and runs locally.

Related event: PrismML's Ternary Bonsai 2 27B Runs at 98.2% Performance in Just 5.9GB(17 posts)→

Original post →

More from Infra

Infra channel →