Alibaba releases Qwen3.8-Flash-Next: 90% training cost cut via new architecture

AI寒武纪 · wechat · 2026-08-26

Alibaba released Qwen3.8-Flash-Next, an early preview of the Qwen4 architecture, as an open-weight multimodal MoE model. It features a 125B backbone with 51B N-gram embedding parameters. Training costs are reduced to approximately one-ninth of Qwen3.7-Plus, while surpassing it in coding and office tasks.

Core Architecture Upgrades

Performance & Pricing

(Note: Claims regarding Zhipu's GLM-5.3-Flash mentioned in the source text are unverified and treated as misinformation.)

Related event: Alibaba Releases Qwen3.8-Flash Multimodal MoE Model to Wide Acclaim(26 posts)→

Original post →

More from Models

Models channel →