Alibaba Launches 2.4T Parameter Qwen3.8-Max with 1M Context and Multimodal Support
togethercompute · x · 2026-08-13
Together AI announced the deployment of Tongyi Qianwen's latest flagship model, Qwen3.8-Max (Qwen3.8-2.4T-A95B). Featuring a sparse Mixture-of-Experts architecture with 2.4 trillion total parameters, it is the largest model released by the Qwen team to date and their first multimodal model exceeding one trillion parameters.
Key Features:
- Massive Context: Natively supports a 1 million token input window with up to 128K output tokens, ideal for processing entire codebases or extensive documents.
- Multimodal & Reasoning: Processes text and image inputs. Always-on thinking mode offers three adjustable reasoning effort levels (low, high, xhigh).
- Agentic Workflows: Optimized for coding and long-horizon agentic tasks, supporting task state dispatch, self-testing (unit/E2E), and structured outputs across loops.
The model was officially announced at the World AI Conference (WAIC) in Shanghai on July 19, 2026, with Alibaba committing to an open-weight release.
Related event: Alibaba Releases 2.4T-Parameter Qwen3.8-Max, Tops HF Trending(15 posts)→
More from coding & agent
- Grok Bot Introduces Zero-Barrier Coding Agent with No Model Selector — ccerrato147 · 2026-08-13
- Boosting Dev Workflow in the Agent Era with Xcode Ad-Hoc Builds — rounak · 2026-08-13
- ExcaliDash: A Self-Hosted Dashboard for Excalidraw with Live Collaboration — tom_doerr · 2026-08-13
- Omarchy Leans Fully Into the Age of AI Agents to Democratize Linux — dosco · 2026-08-13
- Qwen-CUA Launches Native Computer-Use Agent with Red Team Safety Benchmark — hhsun1 · 2026-08-13
- Qwen3.8-2.4T-A95B for long-running agents: 256K context, self-testing, function calling — togethercompute · 2026-08-13