Kimi K3 plus Sol routing lifts solve rate to 85.6% at $7.30 per task
zainhas · x · 2026-07-23
A routing strategy between Kimi K3 and Sol can beat either model alone if the router is good enough.
The cited result shows that using Kimi K3 first and escalating to Sol only when a verifier rejects the answer reaches 85.6% solved at $7.30/task.
Compared with Sol alone:
- coverage improves by about 13 points
- cost drops by $1.07
The post argues the two models behave differently enough that routing or cascading between them is worthwhile.
Related event: Kimi K3 Max vs GPT-5.6 Sol Max: A Routing Problem in Coding(10 posts)→
More from Models
- Same Echo Maze prompt, three frontier models: all passed visually but shipped the same hidden bug — eyishazyer · 2026-09-11
- Benchmark scores drop from 89% to 19% on new evals — how benchmaxxing breaks leaderboard trust — airesearch12 · 2026-09-11
- ChatGPT tells user their question is too hard and to 'accept dumber answers' — phido3000 · 2026-09-11
- Claude is no longer available for minors as Anthropic rolls out age assurance — Muhammad523 · 2026-09-11
- Developer Building a Unified Leaderboard of All Model Benchmark Scores — airesearch12 · 2026-09-11
- Rumor claims Kimi faked performance by serving Claude; DeepSeek new model surprises in evals — realsohamparekh · 2026-09-11