Bindu Reddy says Kimi K3 is cheap, strong on long tasks, but not frontier
bindureddy · x · 2026-07-24
Bindu Reddy jokes that the US government has become a new authority for LLM benchmarks, after it said Kimi K3 is much worse than US frontier models.
- He says LiveBench AI had already reached the same conclusion a few days earlier.
- In his view, Kimi is a very cheap model in the Sonnet/Opus 4.6 class for long-running tasks.
- But he argues that this does not make it frontier intelligence.
More from Models
- Grok 4.5 rolls out across web, iOS and Android with four modes — XFreeze · 2026-07-24
- GPT-5.6 deletes an Act 1 boss in two turns on a live stream — Jsevillamol · 2026-07-24
- Kimi K3 lands on Together AI at launch for coding and agent workloads — togethercompute · 2026-07-24
- User says ChatGPT 5.6 Pro helped disprove a 22-year-old graph theory conjecture — basedjensen · 2026-07-24
- Qwen 3.6 35B MoE runs on a Xiaomi 12 Pro with 12GB RAM at 2.4 tok/s — Aromatic_Ad_7557 · 2026-07-24
- Kimi K3 Model Arrives on Together AI on Day Zero, July 27 — togethercompute · 2026-07-24