Kimi K3 Ties for 5th in Coding, Beats GPT-5.6 in Value
scaling01 · x · 2026-07-18
In the Artificial Analysis Coding Agent Index, Kimi K3 scored 57, tying for 5th place with models like GPT-5.6 Terra and leading Opus 4.8. In specific evaluations, it delivered outstanding results in Terminal-Bench, DeepSWE, and others. Furthermore, the average cost per task for K3 is only $3.18—55% cheaper than GPT-5.6 Sol max—demonstrating exceptional cost-effectiveness for frontier coding.
More from Models
- K3 seems faster on the biggest coding plan than through OpenRouter, user says — xeophon · 2026-07-21
- Andrew Ng-style distillation joke turns model reuse into a theft punchline — pmddomingos · 2026-07-21
- Grok’s Twitter/X search quality appears to have regressed — ivan_bezdomny · 2026-07-21
- SuperGrok regresses on a simple X-account search task, user says — ivan_bezdomny · 2026-07-21
- Chinese LLM vendors push API prices lower as competition intensifies — sen_o · 2026-07-21
- A prompt appears to leak Fable’s chain of thought in a math-heavy trace — teortaxesTex · 2026-07-21