PrismML's Ternary Bonsai 2 27B is 9x smaller at 5.9GB while keeping 98.2% benchmark performance
alexcovo_eth · x · 2026-09-19
PrismML released Ternary Bonsai 2 27B, a ternary-quantized version of Qwen3.8 27B that is 9x smaller at 5.9GB while retaining 98.2% of the full-precision model's aggregate benchmark performance.
The biggest change over the first Bonsai 27B released two months ago is quality: same footprint, but a materially narrower gap to full precision, with strong gains in agentic coding, multimodal reasoning, and long-horizon tool use. Available today under Apache 2.0 and runs locally.
Related event: PrismML's Ternary Bonsai 2 27B Runs at 98.2% Performance in Just 5.9GB(17 posts)→
More from Infra
- Polymarket puts 23% odds on an orbital AI data center launching by end of 2027 — Polymarket · 2026-09-19
- Phylo cuts AI inference cost 60% with open-weight models on Fireworks as usage doubles monthly — sophiamyang · 2026-09-19
- Only 3% of US engineering students go into chips as AI lures the rest — ns123abc · 2026-09-19
- Cerebras launches Money Agent, a finance assistant powered by Qwen3 27B — irinarish · 2026-09-19
- Cloudflare saved another 100TB of RAM by reworking consistent hashing in Rust — Cloudflare Blog · 2026-09-19
- Magnitude: free open-source desktop engine profiles your hardware, picks and tunes local models — nickbaumann_ · 2026-09-19