Ternary Bonsai 2 27B: 9x smaller, 98.2% of full-precision performance, Apache 2.0
cephaloform · x · 2026-09-18
PrismML announced Ternary Bonsai 2 27B, based on Qwen3.8 27B. It is 9x smaller than its full-precision counterpart while retaining 98.2% of aggregate benchmark performance, with strong gains in agentic coding, multimodal reasoning and long-horizon tool use. Footprint stays at 5.9 GB; released under Apache 2.0.
Related event: PrismML Releases Ternary Bonsai 2 27B: 5.95GB Model Runs in Browser(7 posts)→
More from Infra
- 600 tok/s single-request on Qwen 35B with Ninfer on an RTX Pro 6000 — CharlesStross · 2026-09-18
- Cadence sees India's EDA market doubling to $7.82B by 2031 — bookwormengr · 2026-09-18
- Google Open-Sources Agent Substrate on GKE: 10x Density, 1,000+ Dormant Agents per Host — blaizedsouza · 2026-09-18
- Redditor crams six V100 GPUs into a standard full-tower case for local LLM inference — Odd_Caterpillar_2994 · 2026-09-18
- Crusoe raises $3.9B at $30.9B valuation to build data centers and modular AI factories — TechCrunch AI · 2026-09-18
- A 2.5-hour first-principles primer on the semiconductor supply chain worth your time — blaizedsouza · 2026-09-18