Alibaba Releases Qwen3.8-Max: 2.4T Parameter MoE with Native Multimodal Support
baseten · x · 2026-08-13
Alibaba has officially released its latest open-weight frontier reasoning model, Qwen3.8-Max, now available day 0 on the Baseten platform.
Key Specs & Features:
- Architecture: A Mixture-of-Experts (MoE) model with 2.4 trillion total parameters and 95B active parameters per token.
- Context & Modality: Natively supports a 1 million token context window and multimodal inputs.
- Target Use Cases: Built for complex workflows and agentic tasks, excelling at reasoning across massive volumes of professional and scientific literature.
Benchmark Performance:
The model scores 58 on the Artificial Analysis Intelligence Index (up from 47 in its predecessor), ranking #9 across 185 benchmarked models.
Related event: Alibaba Open-Sources 2.4T Parameter Flagship Qwen3.8-Max(17 posts)→
More from Models
- DeepSeek Cybersecurity Test: Top Recall but Bottom Precision — teortaxesTex · 2026-08-13
- Building 'The Office' Agent Simulation with Grok 4.6: A Major Leap in Speed and Capability — mattyp · 2026-08-13
- Sakana AI Updates Chat with New Fugu Model and Code Execution — SakanaAILabs · 2026-08-13
- Gemini V4-Pro Disappoints in Tests, Suspected to Be Hampered by Internal Distillation — teortaxesTex · 2026-08-13
- DeepSeek V4-Pro Ranks #2 Open-Weight Model, Accused of Relying on pass@2 — teortaxesTex · 2026-08-13
- GPT Models Struggle with Pixel Art: UI Interaction Fails and Poor Generation — breath_mirror · 2026-08-13