PrismML's ternary-compressed Bonsai 2 27B lands on OpenRouter at just ~8.5GB

gajesh · x · 2026-09-19

PrismML's reasoning model Ternary Bonsai 2 27B is now live on OpenRouter. Derived from Qwen3.8-27B, it supports coding, math, tool calling, and image understanding with a 262K context window, thinking by default at xhigh reasoning effort.

Its key feature is ternary compression: language-model weights shrink to roughly 8.5GB while retaining 98.2% of the base model's average score across PrismML's 14 thinking-mode benchmarks, enabling efficient inference on consumer hardware. Pricing is $0.075/$0.50 per 1M input/output tokens.

Related event: PrismML's Ternary Bonsai 2 27B Shrinks to 5.9GB, Retains 98.2% Performance(20 posts)→

Original post →

More from Infra

Infra channel →