Alibaba's Qwen3.8-Flash-Next beats rivals at a fraction of the cost
The Decoder · rss · 2026-08-26
Alibaba releases Qwen3.8-Flash-Next, previewing the Qwen4 architecture. This MoE model activates only 6B of 125B parameters per token. At one-ninth the training cost, it beats larger rivals like DeepSeek-V4-Flash and Claude Opus 4.6 on benchmarks, adding pricing pressure to competitors.
More from Models
- Zhipu GLM-5.3-Flash: Matches Opus 4.8 at 1/40 the Cost, Powered by Domestic Chips — vista8 · 2026-08-27
- TokenSpeed adds Day-0 support for Qwen 3.8 Flash Next architecture — Alibaba_Qwen · 2026-08-27
- Zhipu GLM-5.3 open weights releasing in 22 hours — Yuchenj_UW · 2026-08-27
- AI models show more creativity when talking to each other than in assistant persona — nabeelqu · 2026-08-27
- OpenRouter leaderboard: Real token consumption data outweighs media hype — sujingshen · 2026-08-27
- Qwen 3.8-Next Released with Detailed Technical Report on Architecture — nrehiew_ · 2026-08-27