Qwen3.8-Max Preview Tested: Strong Coding but Slow Thinking
Alibaba released Qwen3.8-Max-Preview, a new flagship base model reportedly featuring 2.4 trillion parameters. Open for testing on the official website and the coding app Qoder, it has triggered extensive hands-on tests within the community, though feedback is notably polarized.
Highlights: Coding and Workflows
The model received praise for programming and task execution. @johnseach claimed it wrote 1,500 lines of astrophysics code in one go with zero errors, while @APPSO found its webpage generation extremely fast, capable of producing complex Three.js 3D pages. In Qoder workflow tests, @HeyNayeem noted the model continuously drives tasks to completion with minimal human intervention. Additionally, @青稞AI integrated it to develop complex features like shared wallets, and @Askmasrmod praised its outstanding creative writing capabilities.
Controversies: Slow Thinking and Unmet Expectations
Despite its impressive capabilities, the preview version has clear flaws and divided口碑. @curiousily found the model sometimes gets stuck in a thinking loop, with front-end skills falling short of promotional claims. @Askmasrmod pointed out that its biggest issue is excessively long thinking times, even for simple prompts. Overall, @davidtsong observed sharply divided feedback on X, with some arguing it fails to meet the expected "Fable level." @teortaxesTex candidly stated that Qwen preview versions are typically rough, serving more as scientific references, and estimated its true performance around the GLM 5.2 tier or slightly better.
Testing Suggestions
Addressing the model's instability, @terryyuezhuo (repost) suggested evaluating the API rather than the web UI, as the web chat is merely a free trial portal with inconsistent output quality. @vista8 also reminded users that the web interface is currently best suited for text and simple code tests, though friends reported the model is indeed getting stronger.
2026-07-19 ~ 2026-07-21 · 13 related posts
- Episode 1: Alibaba Announces 2.4T Open-Weight Model Qwen3.8(2026-07-19, 22 posts)
- Episode 2: Leaked Alibaba Qwen 3.8 Max Shows Strong Benchmark and Coding Performance(2026-07-19, 6 posts)
- Episode 3: Qwen3.8-Max-Preview Rolls Out Across Web, PC and iOS Preview(2026-07-19, 2 posts)
- Episode 4: Qwen3.8-Max Preview Tested: Strong Coding but Slow Thinking(2026-07-19, 13 posts)
- Episode 5: Alibaba's Qwen3.8-Max-Preview iterates daily with improved frontend capabilities(2026-07-20, 5 posts)
- Episode 6: Alibaba Announces Open-Weight Qwen3.8 and Multiple New Updates(2026-07-23, 2 posts)
- Episode 7: Alibaba releases Qwen3.8-Max, open-sources weights next week(2026-08-03, 44 posts)
- Episode 8: Qwen3.8-Max Ranks Top Tier Across Benchmarks, Open-Source Narrows Gap(2026-08-03, 7 posts)
- Episode 9: Rumor: Alibaba's Qwen3.8-Max Outperforms Fable 5 and Set to Open Source(2026-08-03, 2 posts)
- Episode 10: Qwen3.8-Max Initial Tests Show Performance on Par with DeepSeek(2026-08-03, 2 posts)
- Episode 11: Alibaba's Qwen3.8-Max: Open-Source Model Nears Closed-Source Frontier(2026-08-03, 13 posts)
- Episode 12: Kimi K3 and Qwen3.8 Max Evaluations Approach Top Closed-Source Models(2026-08-05, 2 posts)
- Episode 13: Rumored Alibaba Qwen3.8-Max to Feature 2.4T Parameters with 95B Active(2026-08-05, 5 posts)
- Episode 14: Alibaba's Qwen3.8-Max tops agentic benchmark, open-sources next week(2026-08-06, 11 posts)
- Episode 15: Testing Alibaba's Qwen3.8-Max: Stellar Long-Context Agent Capabilities(2026-08-06, 3 posts)
- Episode 16: Alibaba's Qwen 3.8 Models Rumored for Next Week Release(2026-08-07, 2 posts)
- Episode 17: Rumors Claim Alibaba Released Qwen3.8-Max(2026-08-09, 2 posts)
Primary sources
- Qwen3.8-Max Preview Tested & The Rise of Chinese Open-Source LLMs — APPSO ·
- Hands-on with Qwen 3.8: Prone to Thinking Loops — curiousily_ ·
- Qwen3.8 Workflow Performance in Qoder — HeyNayeem ·
- Qwen3.8-Max-Preview Praised for Coding Performance — johnseach · 2026-07-19
- [source] Qwen3.8 Workflow Performance in Qoder — HeyNayeem · 2026-07-19
- Qoder Can Finish Tasks Autonomously — HeyNayeem · 2026-07-19
- Qwen3.8 Preview Hands-On Test — 青稞AI · 2026-07-20
- [source] Hands-on with Qwen 3.8: Prone to Thinking Loops — curiousily_ · 2026-07-20
- Conservative Expectations for Qwen 3.8 Preview — teortaxesTex · 2026-07-20
- Can Qwen 3.8 Preview Approach the Frontier? — teortaxesTex · 2026-07-20
- Hands-On Review and Benchmark: Qwen 3.8 Max — WorldofAI · 2026-07-20
- [source] Qwen3.8-Max Preview Tested & The Rise of Chinese Open-Source LLMs — APPSO · 2026-07-20
- Evaluate Qwen 3.8 Max Preview via API, not the chat website — terryyuezhuo · 2026-07-20
- Early Twitter sentiment on Qwen 3.8 looks mixed — davidtsong · 2026-07-21
- Qwen 3.8 Max wins on writing but wastes minutes in thinking mode — Askmasr_mod · 2026-07-21
- Qwen3.8-max-Preview can be tested directly in the browser, with users reporting stronger code generation — vista8 · 2026-07-21