Ternary Bonsai 2 hits Hugging Face: 27B ternary reasoning model under 6GB, runs in browser
cephaloform · x · 2026-09-18
Ternary has released Bonsai 2 on Hugging Face — a 27B reasoning model using ternary weights, built for the agentic era.
Key numbers:
- 9x smaller than FP16 while retaining 98.2% of the intelligence
- Under 6GB, runs 100% locally in the browser via WebGPU
- Available now for anyone to try
A notable showcase of extreme low-bit quantization for local and in-browser deployment.
Related event: PrismML Releases Ternary Bonsai 2 27B: 5.95GB Model Runs in Browser(7 posts)→
More from Infra
- 600 tok/s single-request on Qwen 35B with Ninfer on an RTX Pro 6000 — CharlesStross · 2026-09-18
- Cadence sees India's EDA market doubling to $7.82B by 2031 — bookwormengr · 2026-09-18
- Google Open-Sources Agent Substrate on GKE: 10x Density, 1,000+ Dormant Agents per Host — blaizedsouza · 2026-09-18
- Redditor crams six V100 GPUs into a standard full-tower case for local LLM inference — Odd_Caterpillar_2994 · 2026-09-18
- Crusoe raises $3.9B at $30.9B valuation to build data centers and modular AI factories — TechCrunch AI · 2026-09-18
- A 2.5-hour first-principles primer on the semiconductor supply chain worth your time — blaizedsouza · 2026-09-18