Alibaba Releases Qwen3.8-Max: 2.4T Parameters and 1M Context
FellMentKE · x · 2026-08-06
Alibaba and the Qwen team have introduced Qwen3.8-Max, their most capable flagship model to date.
- Architecture & Scale: Built on a cutting-edge Sparse MoE architecture, it features 2.4T total parameters but activates just 95B for massive efficiency.
- Context & Modality: Natively multimodal across vision and text, it supports a 1M token context window, capable of processing massive inputs like 200+ page documents or 100-hour streams.
- Agentic Capabilities: Designed for long-horizon task execution, it reportedly completed a 16-day real-world software engineering project independently and outperformed humans in the WWW2025 Multimodal Dialogue Intent Recognition Challenge.
More from Models
- Hark Launches Handoff Model, Claims to Outperform ChatGPT 5.4 and Opus 4.8 — peterjliu · 2026-08-06
- Developers Shift to Kimi K3 Amid Anthropic's Limits and Restrictions — casper_hansen_ · 2026-08-06
- DeepSeek Flash v4 Back Online with Significant Speed Improvements — bindureddy · 2026-08-06
- ChatGPT-5.6-luna Cuts Routine Text Processing Costs by 3-5x vs Competitors — zakkohane · 2026-08-06
- Alibaba Releases Qwen3.8-Max with Native Multimodal Reasoning — FellMentKE · 2026-08-06
- Developer Slams GPT 5.6 for Chaotic Coding: Generates 30+ Files But Fails Basic Writing — obinopaul · 2026-08-06