Kimi K3 Ranks 2nd in Coding Test, Distillation Praised

teortaxesTex · x · 2026-07-19

ValsAI's internal benchmark, Vibe Code Bench, shows the Kimi K3 model ranking second overall with a score of 85.0%. The test primarily evaluates a model's ability to create web applications from scratch.

Commenter teortaxesTex expressed amazement, praising the Kimi team's excellent work on model distillation. They noted it is even closer to the Fable model's level than Anthropic's Sonnet 3.5, implying that US companies' legal compliance concerns might be limiting their models' performance ceilings.

Original post →

More from Models

Models channel →