GLM-5.3 hits frontier range; Cerebras claims 30x GPU inference speed

创业邦 · wechat · 2026-08-20

Key items: Korean startup Motif trained Motif3 on $25M of government-backed Nvidia B200 GPUs, scoring 47 on the Artificial Analysis Intelligence Index—level with Alibaba's Qwen3.7 Max—while exploring a token-as-a-service model. Zhipu's GLM-5.3 API went live, scoring 60 on the AA index to enter the global frontier range, tied with Kimi K3 as top open model; pricing unchanged, weights open-sourcing next Friday. Cerebras unveiled CS-4 with three WSE-3 Turbo wafer-scale engines at 750 PFLOPS, claiming 4,400+ tokens/sec/user on GPT-OSS-120B—up to 30x GPU solutions—shipping Q3.

Also: OpenAI's Greg Brockman said the largest planned frontier RL training run is paused for two weeks to strengthen safety monitoring; multiple US states (PA, NY, TX) are tightening data center approvals; Google Cloud says it is automating part of its forward-deployed engineers' enterprise data governance work with AI.

Original post →

More from Models

Models channel →