GLM-5.3 Flash gives 6x more tokens per dollar than Gemini 3.8 Flash in real test
MaziyarPanahi · x · 2026-09-04
A developer measured "intelligence per dollar" on OpenRouter: GLM-5.3 Flash offers 30M tokens for $1 with an Intelligence Index of 57, while Gemini 3.8 Flash scores 59 but only 5M tokens for the same dollar. Same cap, 96% cached. His takeaway: GLM trails by 2 points yet delivers 6× the tokens — "cheap intelligence is getting ridiculous."
Related event: GLM-5.3 Flash Delivers 6x More Tokens Per Dollar Than Gemini(3 posts)→
More from Models
- New local LLM benchmark tracks prefill speed from RTX 5090 down to Raspberry Pi — maximelabonne · 2026-09-04
- Sakana AI's Takuya Akiba to unpack Kimi K3's architecture: how a 2.8T-param open model was built — tkasasagi · 2026-09-04
- Small model Luna praised for beating DeepSeek and its uptime for personal agents — bindureddy · 2026-09-04
- GLM-5.3 gets updated chat template: tool-result reordering now exits early — victormustar · 2026-09-04
- Qwopus 3.8 27B Flash fine-tune ships: 12.8% faster decoding, 80.7% MTP acceptance on Qwen3.8-27B — EAccelerate_42 · 2026-09-04
- Gemini 3.8 Flash edges out Astra on DeepSWE: 73.8% vs 73.3% — jon_barron · 2026-09-04