Qwen3.8 2.4T MoE FP8 Quantization Hits Hugging Face Trending
Qwen · hf · 2026-08-13
Alibaba's Qwen3.8-2.4T-A95B model and its FP8 quantized version recently trended on Hugging Face.
The model features a massive 2.4T total parameters using a Mixture-of-Experts (MoE) architecture, with around 95B active parameters. This release utilizes FP8 precision and is compatible with transformers and various inference endpoints.
Related event: Alibaba's Qwen3.8-Max 2.4T MoE Model Tops Hugging Face Trending(11 posts)→
More from Models
- Qwen 1-bit Quantization Shrinks Model to 397GB, a 91% Reduction — danielhanchen · 2026-08-13
- Users Complain About Claude's Unstoppable Chain-of-Thought Output — GuyHachmon · 2026-08-13
- OpenAI's gpt-live-1 Achieves Near-Perfect Conversational Turn Detection — pbbakkum · 2026-08-13
- xAI Launches Grok 4.6 Across Cursor, API with 2x Token Promo — aman_madaan · 2026-08-13
- xAI Exec Hints at Grok 4.5: Focused on Coding Agents — aman_madaan · 2026-08-13
- Reddit Speculates: Is Grok 4.6 a Fine-tune of Kimi K3? — robertpro01 · 2026-08-13