Alibaba Upgrades Qwen3.8-Max-0902: 2.4T Params, 1M Context, at $2/$6 per Million Tokens
thione · x · 2026-09-07
Alibaba Cloud upgraded Qwen3.8-Max-0902: 2.4T parameters with a 1M-token context window, further post-trained on Coding & Cowork for stronger performance on complex enterprise tasks, scientific research, and long-horizon workflows.
Pricing is $2 input / $6 output per million tokens, with explicit cache hits at $0.17 and implicit cache hits at $0.25. Available via Model Studio and Qwen Cloud APIs.
More from Models
- User notes Astra follows strict skill rules less reliably than Sol — petergyang · 2026-09-07
- Astra still generates fake unrelated criticisms when fact-checking, user finds — AndyMasley · 2026-09-07
- MLP paper shows neurons turn monosemantic in clustered regression, challenging global subspace view — burkov · 2026-09-07
- Early user: OpenAI's Astra is faster, more token-efficient and higher quality on hard tasks — Yamapama · 2026-09-07
- Gemini 4 Deep Think checkpoint tested: 85k-token SVG generations — Ryoiki-Tokuiten · 2026-09-07
- Gemini's agentic video understanding cuts tokens by 88% and costs by 66% — patloeber · 2026-09-07