Qwen's New 2.4T MoE Model Tops Hugging Face Trending
Qwen · hf · 2026-08-12
Alibaba's latest Qwen model, Qwen/Qwen3.8-2.4T-A95B, has surged to the top of the Hugging Face trending list.
The model utilizes a Mixture-of-Experts (MoE) architecture with a massive total parameter count of 2.4T and 95B active parameters. It supports text-generation pipelines, is endpoints compatible, and is designed for conversational and text generation tasks.
Related event: Alibaba Releases Qwen3.8-2.4T-A95B Model(7 posts)→
More from Models
- Liquid AI Launches LFM2.5-VL-3B: A Lightweight Vision-Language Model Outperforming 2.6x Larger Rivals — JosephJacks_ · 2026-08-13
- Grok Offers 85% Discount Over OpenAI with Similar Performance — GavinSBaker · 2026-08-13
- Do LoRAs Fail to Work on Pruned MiniMax H3 Models? — kayteee1995 · 2026-08-13
- LiquidAI Launches 3B Vision-Language Model LFM2.5-VL, Outscoring Larger Rivals — JosephJacks_ · 2026-08-13
- 21-Year-Old Math Enigma Solved by Human; GPT and Claude Both Failed — anshulkundaje · 2026-08-13
- Reviewing AI Like an Art Critic: Grok 4.6 Tested on Astrology & Philosophy — karinanguyen · 2026-08-13