Alibaba Open Sources Qwen3.8-Flash-Next with DeepSeek-Inspired Architecture

量子位 · wechat · 2026-08-27

Alibaba released the weights for Qwen3.8-Flash-Next, a preview of the Qwen4 architecture. The 125B MoE model features 51B N-gram Embedding parameters (inspired by DeepSeek's Engram), significantly reducing training costs to 1/9th while improving performance. It also upgrades attention (QSA), residual connections, and the optimizer (Muon). The model supports 1M token context, offers API costs roughly 1/12th of the Max version, and is compatible with OpenAI/Anthropic protocols.

Related event: Alibaba Open-Sources Qwen3.8-Flash-Next, A 6B-Activated Preview of Qwen4 Architecture(18 posts)→

Original post →

More from Models

Models channel →