Alibaba releases Qwen3.8 Max: 2.4T-param MoE, 1M context, open for commercial use
AdinaYakup · x · 2026-08-13
Qwen3.8 Max, the largest model in the Qwen family, is now available on Hugging Face. It features a 2.4T total parameter MoE with 95B active parameters, supports 1M context, and offers an FP8 version. It combines Gated DeltaNet and Gated Attention for better reasoning and efficiency. The license is open for developers and companies to use, fine-tune, and commercialize, with extra licensing only for very large-scale AI businesses.
Related event: Alibaba's Qwen3.8-Max 2.4T MoE Model Tops Hugging Face Trending(11 posts)→
More from Models
- Qwen 1-bit Quantization Shrinks Model to 397GB, a 91% Reduction — danielhanchen · 2026-08-13
- Users Complain About Claude's Unstoppable Chain-of-Thought Output — GuyHachmon · 2026-08-13
- OpenAI's gpt-live-1 Achieves Near-Perfect Conversational Turn Detection — pbbakkum · 2026-08-13
- xAI Launches Grok 4.6 Across Cursor, API with 2x Token Promo — aman_madaan · 2026-08-13
- xAI Exec Hints at Grok 4.5: Focused on Coding Agents — aman_madaan · 2026-08-13
- Reddit Speculates: Is Grok 4.6 a Fine-tune of Kimi K3? — robertpro01 · 2026-08-13