ZAI Releases GLM-5.3-Flash on Chinese Chips
AccBalanced · x · 2026-08-27
ZAI announced GLM-5.3-Flash (formerly Ox Alpha), a 320B MoE model released under the MIT license. It features a 1M-token context window and native multimodal capabilities. Notably, it runs entirely on Chinese AI chips (likely Ascend 910C), demonstrating an engineering breakthrough for domestic compute infrastructure.
Related event: Zhipu Open-Sources GLM-5.3-Flash: Frontier Intelligence at Low Cost(39 posts)→
More from Infra
- vLLM hits 130k tok/s on DeepSeek V4 Pro in AgentX benchmark — AccBalanced · 2026-08-27
- Google Cloud Run Introduces Instances for MicroVM Deployment — steren · 2026-08-27
- Anthropic Signs $4.5B Compute Deal for Nvidia Rubin Chips at Nscale — Beth_Kindig · 2026-08-27
- DeepSeek-v4-Pro Generates 130M Tokens for $1, 77x Cheaper Than Opus — bookwormengr · 2026-08-27
- NVIDIA FLARE Cuts Federated VLM Training Traffic by 99% — dl_weekly · 2026-08-27
- Same Budget: 256GB Mac or Two DGX Sparks for 70B Inference? — Whyme-__- · 2026-08-27