Kimi K3 Security Eval: High Exploit, Low Guardrails

Evaluations show Kimi K3 can autonomously develop exploits and outperform GLM-5.2, but its end-to-end attack capabilities lag behind US frontier models. Its guardrails also failed to block malicious instructions, unlike OpenAI and Anthropic models.

2026-07-26 ~ 2026-07-28 · 4 related posts

Full story(2 episodes)→