Bonsai 2 tested: 98% of Qwen 27B in 6GB, but zero working agent builds

Prompt Engineering · youtube · 2026-09-28

The Prompt Engineering YouTube channel tested PrismML's Bonsai 2, a ternary compression scheme claiming to keep 98% of the Qwen 27B model's performance in a 6GB file.

Findings:

Takeaway: compressed models may be fine for chat, but think twice before handing them long-horizon agent tasks.

Related event: Bonsai 2 Compresses Qwen 27B to 6GB but Falters on Long Tasks(2 posts)→

Original post →

More from Models

Models channel →