Qwen and GLM release new MoE models focusing on low cost and high performance
thione · x · 2026-08-31
Alibaba's Qwen released Qwen3.8-Flash-Next, an open-weight multimodal Mixture-of-Experts model designed for significantly lower training and inference costs. Simultaneously, Zhipu GLM released open-weight GLM-5.3-Flash, a 320B-parameter MoE model aimed at achieving frontier coding performance at low cost.
Related event: Qwen and Zhipu Release New Open-Weight MoE Models(2 posts)→
More from Models
- Tiel-Coder-35B-A3B Trends on HF with Speculative Decoding & MTP — peculiar-ragdoll · 2026-08-31
- Google Releases Gemini Omni 1.1 Flash, Updating Its Fast Multimodal Model for Developers — thione · 2026-08-31
- DeepSeek launches low-cost vision model; Anthropic previews hardware control protocol for agents — thione · 2026-08-31
- Professor says Claude grammar fixes got his writing flagged 100% AI by Pangram — ipeirotis · 2026-08-31
- Anthropic: Claude-powered automated alignment researchers beat veteran humans' ideas — burny_tech · 2026-08-31
- Anthropic usage tiers offer only 3.5x-6x limits, far below advertised — RichmanRonald · 2026-08-31