GLM-5.3 hits frontier range; Cerebras claims 30x GPU inference speed
创业邦 · wechat · 2026-08-20
Key items: Korean startup Motif trained Motif3 on $25M of government-backed Nvidia B200 GPUs, scoring 47 on the Artificial Analysis Intelligence Index—level with Alibaba's Qwen3.7 Max—while exploring a token-as-a-service model. Zhipu's GLM-5.3 API went live, scoring 60 on the AA index to enter the global frontier range, tied with Kimi K3 as top open model; pricing unchanged, weights open-sourcing next Friday. Cerebras unveiled CS-4 with three WSE-3 Turbo wafer-scale engines at 750 PFLOPS, claiming 4,400+ tokens/sec/user on GPT-OSS-120B—up to 30x GPU solutions—shipping Q3.
Also: OpenAI's Greg Brockman said the largest planned frontier RL training run is paused for two weeks to strengthen safety monitoring; multiple US states (PA, NY, TX) are tightening data center approvals; Google Cloud says it is automating part of its forward-deployed engineers' enterprise data governance work with AI.
More from Models
- Agent cost-efficiency plot: Qwen hits Pareto frontier — MikePFrank · 2026-08-20
- Analysis estimates GPT-5.6-Sol params at 1.5T to 2.5T — scaling01 · 2026-08-20
- ChatGPT Down: Login and Signup Issues Reported — ns123abc · 2026-08-20
- Models can now code entire complex software in under an hour — BLUECOW009 · 2026-08-20
- Upcoming benchmark: Local deployment comparison of Qwen3.8, Gemma4, and GPT-OSS — karminski3 · 2026-08-20
- Claude Code Faces Rough Month; Alternative Models Evaluated — Hesamation · 2026-08-20