Kimi K3 is called roughly equivalent to Opus 4.8 on ALE-Bench

scaling01 · x · 2026-07-22

Kimi K3 is compared to Opus 4.8 on ALE-Bench

A short X post claims Kimi K3 is “basically Opus 4.8” on ALE-Bench, while Inkling and Grok 4.5 are not competitive in that comparison.

The attached chart plots performance against cost and labels several frontier models, including GPT 5.6 Sol, Fable 5, GPT 5.5, Gemini 3.1 Pro, Opus 4.8, Kimi K3, Inkling, and Grok 4.5.

Because the post is a benchmark-style model comparison rather than a product workflow or coding-agent story, it belongs in the models channel.

Original post →

More from Models

Models channel →