Government safety report says Kimi K3 trails U.S. frontier models on cyber offense
rohanpaul_ai · x · 2026-07-24
Government safety institutes benchmark Kimi K3’s cyber skills
CAISI and UK safety institutes published a preliminary evaluation of Kimi K3’s offensive cyber capabilities against frontier U.S. models.
- Kimi K3 stopped at step 17 on average, while the strongest U.S. models reached 28.5.
- On ExploitBench, leading U.S. models scored 76.2%, versus 32.2% for Kimi K3 and 24.4% for GLM-5.2.
- The report says Kimi K3 can autonomously carry out meaningful parts of an attack and sometimes complete an entire simulated enterprise attack, but it still trails frontier U.S. models and does not reliably refuse offensive requests.
- The U.S. models in this comparison were evaluated with safeguards switched off, so the results reflect capability ceilings rather than shipping product behavior.
More from Safety
- Novosad backs Hassabis' AI safety institution-building over kneecapping US labs — paulnovosad · 2026-09-11
- Economist argues safe AGI comes from engineers inside big labs, not regulation — paulnovosad · 2026-09-11
- LLM-driven attacks mostly follow Pentesting 101: traditional defenses still work — AccBalanced · 2026-09-11
- Op-ed: the ">10% extinction" narrative is liability evasion — AI is just software, and the vendor is the defendant — gerardsans · 2026-09-11
- GreyNoise reveals campaign run by hundreds of AI agents against PaperCut NG/MF — AccBalanced · 2026-09-11
- Economist Warns US Collective Action Could 'Regulate AI Progress Out of Existence' — paulnovosad · 2026-09-11