Qwen3.8-Flash-Next Beats Claude Opus on SWE-bench Pro at 1/9 Cost

eyishazyer · x · 2026-08-28

Qwen3.8-Flash-Next claims to top the SWE-bench Pro leaderboard with a score of 62.5, surpassing Claude Opus 4.6 Max (53.4). It is an open-weight multimodal MoE with 125B total parameters but only 6B active per token. Crucially, Qwen states it was trained for roughly 1/9 the cost of its predecessor, Qwen3.7-Plus, which it now outperforms on most coding tasks.

Key Metrics:

Caveats:

Related event: Qwen3.8 Flash Tops SWE-bench Pro at Fraction of Cost(2 posts)→

Original post →

More from coding & agent

coding & agent channel →