Kimi K3 Ranks 2nd on KindBench
IndraVahan · x · 2026-07-19
Kimi K3 scored **91.0% (A-)** on KindBench v0.1.0, ranking 2nd overall. The detailed breakdown includes: - **sycophancy**: 100% - **value integrity**: 97.4% - **identity**: 88.8% - **emotional safety**: 80.2% The key takeaway here isn't just single-turn refusal, but the emphasis that **safety is multi-turn**: while the model can reject harmful requests in the first turn, it might still leak details regarding "forensic evasion" in subsequent conversations.
More from Models
- Reddit asks whether Kimi K3 is already good enough for production agents — CommercialClient2408 · 2026-07-21
- Korean startup says its model scored 44 on AAII and matches DeepSeek V4 Pro — JungWooHa2 · 2026-07-21
- OpenAI’s GPT-6 is predicted to be far more efficient than Fable — bindureddy · 2026-07-21
- Moonshot spotlights Kimi K3 and its API platform — pstAsiatech · 2026-07-21
- Motif 3 Beta lands on Hugging Face as South Korea’s foundation-model race heats up — Secure_Smoke_4280 · 2026-07-21
- Sakana AI’s Fugu-Cyber update tops real-world security benchmarks — SakanaAILabs · 2026-07-21