GLM-5.3-Flash matches Opus 4.8, 45x cheaper
Hesamation · x · 2026-08-26
Evaluations reveal that Zhipu's GLM-5.3 Flash scores 57 on the Artificial Analysis intelligence index, matching Opus 4.8. However, it costs only $0.045 per task, making it 45x cheaper than Opus 4.8 and 15x cheaper than GLM-5.3. It also achieves 3x lower attention compute and 4.4x smaller KV cache at 1M context.
More from Models
- First test of MiniMax H3 video generation: 8s video took 1 hour on rented A100 — Winter_Assignment_78 · 2026-08-27
- Anthropic reportedly releasing Fable 5.1 model soon — mark_k · 2026-08-27
- Observers doubt Simile/Aaru claims over lack of datasets and peer review — daveholtz · 2026-08-27
- Navigator n2 released: A frontier 27B computer-use model — DhruvBatra_ · 2026-08-27
- Alibaba Releases FP8 Quantized Qwen3.8-Flash-Next Model — Qwen · 2026-08-27
- Qwen 3.8 27b coding performance shocks community, rivaling GPT 5.5 on consumer hardware — GrokiniGPT · 2026-08-27