GLM 3.8-27B quality 'absurdly superior' to 35B-A3B models, saving 22-33% tokens

JLeonsarmiento · reddit · 2026-09-12

An applied-science researcher shared a hands-on account of using GLM 3.8-27B (a Qwen-based dense model) for full research workflows, calling it 'absurdly superior' to every 3.5/3.6-35B-A3B variant (kat, Ornith/tiel, nex-2, etc.), with only Ornith coming close.

The author's verdict: accept the slower speed for the quality and let the 'fat bottom Qwen' work.

Original post →

More from Models

Models channel →