Alibaba releases Qwen3.8 Max: 2.4T-param MoE, 1M context, open for commercial use

AdinaYakup · x · 2026-08-13

Qwen3.8 Max, the largest model in the Qwen family, is now available on Hugging Face. It features a 2.4T total parameter MoE with 95B active parameters, supports 1M context, and offers an FP8 version. It combines Gated DeltaNet and Gated Attention for better reasoning and efficiency. The license is open for developers and companies to use, fine-tune, and commercialize, with extra licensing only for very large-scale AI businesses.

Related event: Alibaba's Qwen3.8-Max 2.4T MoE Model Tops Hugging Face Trending(11 posts)→

Original post →

More from Models

Models channel →