Qwen3.8-Max Architecture: 95B Active Parameters Delivers Outperforming Compute Efficiency

thetripathi58 · x · 2026-08-06

Discussing Alibaba's newly released Qwen3.8-Max, this post highlights its Mixture-of-Experts (MoE) architecture, which activates only 95 billion parameters out of a total of 2.4 trillion. The author notes that despite running significantly less compute per request compared to heavier models, real-world testing across finance, web development, and photorealistic rendering shows it surprisingly outperforming models with much higher compute consumption.

Original post →

More from Models

Models channel →