Qwen3.8 Preview Leaks, Open Models Honestly Benchmark Against Fable 5
aigclink · x · 2026-07-19
Alibaba has released a Qwen3.8-Max preview with 2.4T parameters, marking its return to the open-source community, with weights expected to be open-sourced after the official launch. Benchmarks are based on 400 real-world tasks from Alibaba's own agentic products (using isolated ECS and multiple cross-validations), surpassing Kimi K3 and GLM-5.2 in Coding and Cowork. Officially positioned as "second only to Fable 5", the model is now available on platforms like Token Plan and Qoder, with hands-on tests in Cline showing performance close to 4.8.
Amid fierce competition among domestic open-source models, Qwen quickly followed up on Kimi K3's release within 3 days. The author notes that Chinese developers are no longer blindly claiming to be "number one globally"; instead, they openly acknowledge the gap with the strongest closed-source model (Fable 5). This honesty is proving more credible than inflated benchmark scores.
Related event: Alibaba Announces 2.4T Open-Weight Model Qwen3.8(22 posts)→
More from Models
- BullshitBench update: GPT-6-Astra beats all prior OpenAI models but still trails Anthropic — scaling01 · 2026-09-11
- Astra Scores 83% on GauntletBench, First Computer-Use Agent to Beat Human Baseline — ducha_aiki · 2026-09-11
- Kimi K2.8 Preview rolls out: near-K3 coding performance, 1M context for all tiers — teortaxesTex · 2026-09-11
- Looking for a classifier of software engineering task shapes to pick models per task — StewartalsopIII · 2026-09-11
- DeepSeek V4 Pro API to continue after Sept 2026, billing unchanged — teortaxesTex · 2026-09-11
- DeepSeek V4.1 Flash Hits 98% of GPT-6 Astra's Score at 1.4% of the Cost in Third-Party Benchmark — ayushtweetshere · 2026-09-11