Ternary Bonsai 2 27B: 9x smaller than Qwen3.8 27B at 98.2% performance, Apache 2.0
cephaloform · x · 2026-09-18
PrismML released Ternary Bonsai 2 27B, a ternary-quantized version of Qwen3.8 27B that is 9x smaller than its full-precision counterpart, weighs just 5.9 GB, and retains 98.2% of the aggregate benchmark performance. Compared to the first Bonsai 27B from two months ago, the biggest change is quality: gains are strongest in agentic coding, multimodal reasoning, and long-horizon tool use. Released under Apache 2.0, with the team claiming it now closes in on Qwen3.8 27B across coding, math, instruction following, agentic, and vision tasks.
Related event: PrismML Releases Ternary Bonsai 2 27B: 5.95GB Model Runs in Browser(7 posts)→
More from Models
- Leaked GPT-6 Astra tops FrontierSWE at 65.5%, compresses 600MB audio to 20KB — soumitrashukla9 · 2026-09-18
- Muse Code 'crazy cheap and quite good': early take on Meta's coding model — AIandDesign · 2026-09-18
- Community unmasks stealth model Union Alpha as Unbiased's Pareto 26.9 in one day — gaganghotra_ · 2026-09-18
- Emad Mostaque hails an 'Opus 4.5-level' model that runs on 8GB RAM — ccerrato147 · 2026-09-18
- Astra computer use formats Google Docs autonomously, leaving users impressed — var_epsilon · 2026-09-18
- jev may have unseated Cohere as the best reranker on speed, cost, and performance — multiply_matrix · 2026-09-18