Alibaba's Qwen3.8-Flash-Next releasing tomorrow with Qwen4 MoE architecture
rohanpaul_ai · x · 2026-08-25
Alibaba's Qwen team announced the upcoming release of Qwen3.8-Flash-Next tomorrow. Built on the next-gen Qwen4 architecture, the model features a multimodal Mixture-of-Experts (MoE) design with 125 billion total parameters, activating only 6 billion per token for high-efficiency performance.
Related event: Alibaba Teases Qwen3.8-Flash-Next as a Preview of Qwen4 Architecture(6 posts)→
More from Models
- OpenAI's new inference engine boosts throughput up to 4.1x on select models — rohanpaul_ai · 2026-08-25
- OpenAI's Codex lead: next-gen models need more than your laptop, Codex will look primitive — koltregaskes · 2026-08-25
- AI4Bharat's 0.8B IndicOCR Takes On Models 4-5x Its Size, Open Weights — sumanthd17 · 2026-08-25
- Developer Eagerly Anticipates Release of Qwen 3.8 Small Models — chris_j_paxton · 2026-08-25
- AI writing latency stagnant; RL focus may degrade quality — peterwildeford · 2026-08-25
- LLM use is like MCMC: reward-seeking prompts that never backtrack — rickasaurus · 2026-08-25