Rumor: Alibaba's Qwen to release a 125B-A6B MoE model, hailing the era of ~120B models
TheZachMueller · x · 2026-08-25
Developer Zach Mueller quotes a leak that @AlibabaQwen will soon release a 125B total / 6B active MoE model, remarking: "Welcome to the era of 120B models."
The spec continues Qwen's open-source MoE line, trading moderate total parameters for cost-efficient inference. This is an unverified third-party leak.
More from Models
- SenseNova U1.5 quantized to run on 12GB VRAM with INT8/W4A8 — junklont · 2026-08-25
- AI struggles with precise terminology in technical writing — Ben_Reinhardt · 2026-08-25
- Startup Accelerated Understanding launches on neural operators, rumored source of huge-context model — inductionheads · 2026-08-25
- Community speculates on Qwen 4 architecture: Sparse Attention or MLA hybrid? — challis88ocarina · 2026-08-25
- Test of 1,000 prompts: ChatGPT and Google AI share only 2.6% of cited sources — nikvassev · 2026-08-25
- Unsloth AI aims for day-zero llama.cpp support for Qwen models — danielhanchen · 2026-08-25