Benchmark screenshot puts Kimi K3 at 155 and 13th in a 212-model list
iruletheworldmo · x · 2026-07-23
A repost arguing that Kimi is still 7–12 months behind US frontier labs, even before accounting for US government delays in releasing models.
The attached benchmark screenshot shows Kimi K3 with an EPOCH Capabilities Index of 155, ranking 13/212, alongside a cluster of frontier models around it:
- GPT-5.6 Sol: 162
- GPT-5.5 Pro: 161
- Claude Fable 5: 161
- GPT-5.5: 159
- GPT-5.6 Terra / Claude Opus 4.8 / GPT-5.4 Pro: 158
- Kimi K3: 155
The post’s point is not that Kimi is weak, but that the benchmark still places it meaningfully behind the top American models.
Related event: Kimi K3 Sets New Open-Source ECI Record but Still Lags Behind(3 posts)→
More from Models
- Google says Gemini is now routed into Search AI Overviews and AI Mode — gaganghotra_ · 2026-07-23
- Reddit user says Grok beats ChatGPT, Claude and Gemini for long conversations — Astrum10worlds · 2026-07-23
- Reddit asks for hands-on reports on Poolside’s Laguna S 2.1 in agent loops — ForsookComparison · 2026-07-23
- New Open Science Initiative to Build a 1-Trillion-Parameter AI Model — stochasticchasm · 2026-07-23
- Reddit test: ChatGPT, Claude and Gemini all ranked Grok last — soulsintention · 2026-07-23
- More frontier models are coming this week, and Kimi K3 may get buried — bindureddy · 2026-07-23