Kimi K3 Ties for 5th in Coding, Beats GPT-5.6 in Value

scaling01 · x · 2026-07-18

In the Artificial Analysis Coding Agent Index, Kimi K3 scored 57, tying for 5th place with models like GPT-5.6 Terra and leading Opus 4.8. In specific evaluations, it delivered outstanding results in Terminal-Bench, DeepSWE, and others. Furthermore, the average cost per task for K3 is only $3.18—55% cheaper than GPT-5.6 Sol max—demonstrating exceptional cost-effectiveness for frontier coding.

Related event: Moonshot's Kimi K3 Tops Frontend Code Arena, Nearing Fable 5 in Coding at a Third of the Cost(12 posts)→

Original post →

More from Models

Models channel →