Leaked Alibaba Qwen 3.8 Max Shows Strong Benchmark and Coding Performance
Recently, a preview version of Alibaba's Qwen 3.8 Max model was leaked, sparking heated community discussions. The model allegedly has 2.4T parameters and is officially claimed to be the "strongest model except for Fable 5." Currently, it has demonstrated exceptionally strong performance in multiple benchmark and practical development tests.
Benchmarks and Capability Evaluation
Regarding benchmarks, the leaked KingBench 3 leaderboard shows Qwen 3.8 Max scoring 81.25, closely trailing the top-ranked Fable 5 (82.5) and surpassing several Claude models. @bdsqlsz pointed out that in the "candy test," its mathematical ability exceeded GPT 5.5 high and Kimi K3, with outstanding coding capabilities as well. In practical applications, @赛博禅心 tested the model via Claude Code on a development task involving multiple payment logics, confirming its robust coding ability and extremely fast response speed.
Reactions and Impact
Users on platforms like Reddit have been actively verifying the model's authenticity. Faced with the strong benchmark data, @orange stated that if Qwen 3.8 truly surpasses GPT-5.6, the technological gap between Chinese and US models could shorten to about 3 months, though he plans to make a final judgment after actually trying it. Furthermore, @Scobleizer reposted, highlighting a realistic challenge: while open-source frontier models are getting larger, benchmark scores do not equate to affordable inference costs, which is a practical hurdle for the model's future deployment.
2026-07-19 ~ 2026-07-20 · 6 related posts
- Episode 1: Alibaba Announces 2.4T Open-Weight Model Qwen3.8(2026-07-19, 22 posts)
- Episode 2: Leaked Alibaba Qwen 3.8 Max Shows Strong Benchmark and Coding Performance(2026-07-19, 6 posts)
- Episode 3: Qwen3.8-Max-Preview Rolls Out Across Web, PC and iOS Preview(2026-07-19, 2 posts)
- Episode 4: Qwen3.8-Max Preview Tested: Strong Coding but Slow Thinking(2026-07-19, 13 posts)
- Episode 5: Alibaba's Qwen3.8-Max-Preview iterates daily with improved frontend capabilities(2026-07-20, 5 posts)
- Episode 6: Alibaba Announces Open-Weight Qwen3.8 and Multiple New Updates(2026-07-23, 2 posts)
- Episode 7: Alibaba releases Qwen3.8-Max, open-sources weights next week(2026-08-03, 44 posts)
- Episode 8: Qwen3.8-Max Ranks Top Tier Across Benchmarks, Open-Source Narrows Gap(2026-08-03, 7 posts)
- Episode 9: Rumor: Alibaba's Qwen3.8-Max Outperforms Fable 5 and Set to Open Source(2026-08-03, 2 posts)
- Episode 10: Qwen3.8-Max Initial Tests Show Performance on Par with DeepSeek(2026-08-03, 2 posts)
- Episode 11: Alibaba's Qwen3.8-Max: Open-Source Model Nears Closed-Source Frontier(2026-08-03, 13 posts)
- Episode 12: Kimi K3 and Qwen3.8 Max Evaluations Approach Top Closed-Source Models(2026-08-05, 2 posts)
- Episode 13: Rumored Alibaba Qwen3.8-Max to Feature 2.4T Parameters with 95B Active(2026-08-05, 5 posts)
- Episode 14: Alibaba's Qwen3.8-Max tops agentic benchmark, open-sources next week(2026-08-06, 11 posts)
- Episode 15: Testing Alibaba's Qwen3.8-Max: Stellar Long-Context Agent Capabilities(2026-08-06, 3 posts)
- Episode 16: Alibaba's Qwen 3.8 Models Rumored for Next Week Release(2026-08-07, 2 posts)
- Episode 17: Rumors Claim Alibaba Released Qwen3.8-Max(2026-08-09, 2 posts)
Primary sources
- Rumor: Qwen 3.8 Is Coming — OutlandishnessNo5636 · 2026-07-19
- [source] Qwen 3.8 Max Preview Benchmarks Leaked — bdsqlsz · 2026-07-19
- [source] Alleged Qwen3.8 Preview Leaks: 2.4T Parameters, Fast Speed, Strong Coding — 赛博禅心 · 2026-07-19
- [source] Qwen 3.8 Benchmarks Leak Online — mark_k · 2026-07-20
- Qwen 3.8 Rumored to Surpass GPT-5.6 — oran_ge · 2026-07-20
- Qwen3.8 Max reportedly hits 2.4T parameters — Scobleizer · 2026-07-20