Ternary Bonsai 2 27B: 5.9GB quantized model keeps 98.2% performance, runs on Mac

soumitrashukla9 · x · 2026-09-22

PrismML released Ternary Bonsai 2 27B, a ternary-quantized model based on Qwen3 27B that is 9x smaller than its full-precision counterpart, fits in 5.9GB, and retains 98.2% of aggregate benchmark performance, with notable gains in agentic coding, multimodal reasoning and long-horizon tool use; Apache 2.0 licensed. Economist johnjhorton then ran it locally on a 48GB M4 Max MacBook Pro via EDSL and a Metal-enabled llama.cpp fork, executing a battery of economic rationality tests successfully, with code and a write-up open-sourced.

Related event: PrismML Open-Sources Ternary Bonsai 2 27B: 9x Smaller, 98.2% Performance(3 posts)→

Original post →

More from Models

Models channel →