Dev burns 300M tokens on GLM 5.3 in a week and still has quota left
saibharadwaj · x · 2026-09-10
Developer saibharadwaj notes that Chinese models like GLM, DeepSeek, StepFun and MiMo are surprisingly fast, offering huge token limits and bonus tokens with no resets — he consumed about 300M tokens on GLM 5.3 in a week and still had 2% left. He also wonders how Chinese AI firms keep building powerful custom AI chips despite semiconductor export restrictions.
More from Infra
- Analyst: DeepSeek's latest change is a big win for token efficiency, moving toward OpenAI's regime — teortaxesTex · 2026-09-10
- Engram embeddings load overlapped with GPU compute, so fetch time costs nothing — bookwormengr · 2026-09-10
- Mac mini tested: local 35B runtime hits Haiku-level scores but falls short for agents — PawelHuryn · 2026-09-10
- Miles ships Day-0 RL support for DeepSeek-V4.1-Flash with KL held at 0.0012–0.0017 — ying11231 · 2026-09-10
- The data center is a symbol: why debunked claims about AI infrastructure still spread — ShakeelHashim · 2026-09-10
- Kimi K3 lands on RunPod: 2.8T params, 1M context, $3/$15 per 1M tokens — Kimi_Moonshot · 2026-09-10