Qwen3.8 2.4T MoE FP8 Quantization Hits Hugging Face Trending

Qwen · hf · 2026-08-13

Alibaba's Qwen3.8-2.4T-A95B model and its FP8 quantized version recently trended on Hugging Face.

The model features a massive 2.4T total parameters using a Mixture-of-Experts (MoE) architecture, with around 95B active parameters. This release utilizes FP8 precision and is compatible with transformers and various inference endpoints.

Related event: Alibaba's Qwen3.8-Max 2.4T MoE Model Tops Hugging Face Trending(11 posts)→

Original post →

More from Models

Models channel →