New Qwen and GLM Open Models Benchmark Against Claude Opus

kimmonismus · x · 2026-08-26

A significant day for open-weight and local AI. Qwen3.8-Flash-Next (125B MoE, 6B active) beats Claude Opus 4.6 Max on 8 of 9 benchmarks. GLM-5.3-Flash (320B total, 18B active) scores close to Opus 4.8 on Terminal-Bench and leads on several agentic benchmarks. Both models are MIT-licensed, natively multimodal, and support 1M context. While they require serious workstations to run, the narrowing gap with frontier models serves as a wake-up call.

Related event: Z.ai Open-Sources GLM-5.3-Flash: 320B MoE Rivals Claude Opus 4.8 at Fraction of Cost(26 posts)→

Original post →

More from Models

Models channel →