Bonsai 2 QAT models score ~91.5% in independent Qwen3.8 quantization comparison
ali_byteshape · reddit · 2026-09-19
The byteshape team benchmarked Prism-LM's new Bonsai 2 QAT models under a unified Qwen3.8 evaluation methodology, finding 91.5% on their composite benchmark with a strong quality-throughput trade-off, plus separate Instruct and Thinking results.
Related event: Bonsai 2 QAT quantized model scores ~91.5% in independent benchmarks(2 posts)→
More from Models
- Skeptical deep dive confirms Humanity's Last Exam errors; official o3-mini grader marked right answers wrong every time — paul_cal · 2026-09-20
- Matt Shumer asks if Jev could help with scalable oversight and alignment checks — mattshumer_ · 2026-09-20
- FrontierSWE v2 opens 24.1-point gap: Claude Fable 5.1 scores 56.29% vs GPT-5.6's 32.2% — geoffwolfe · 2026-09-20
- 22M local model beats JEV 93% vs 80% on Banking77 in 8ms on CPU — Prompt Engineering · 2026-09-20
- Jev loses to Gemini on 1,565-email classification benchmark, but dev still wants it in production — socialwithaayan · 2026-09-20
- Jev Detector scans ~10,000 words for AI slop in ~2 seconds, free with no sign-up — socialwithaayan · 2026-09-20