CAISI report says Kimi K3 leads open-weight models but trails U.S. frontier systems
mattsheehan88 · x · 2026-07-24
- A Commerce/CAISI-NIST report says the newest Chinese model is still behind top U.S. frontier models, but is the strongest open-weight model.
- The attached charts compare cyber capability and exploit-development performance across models, including Kimi K3 and GLM-5.2.
- In the exploit-development benchmark, top U.S. models score about 76.2%, while Kimi K3 is around 32.2% and GLM-5.2 around 24.4%.
- A second chart places Kimi K3 near the top of the Chinese model cluster on overall cyber capability, though still below the U.S. frontier band.
Related event: CAISI Report: Kimi K3 Leads Open Weights, Lags US Frontier(2 posts)→
More from Models
- GPT-5.6 deletes an Act 1 boss in two turns on a live stream — Jsevillamol · 2026-07-24
- Kimi K3 lands on Together AI at launch for coding and agent workloads — togethercompute · 2026-07-24
- User says ChatGPT 5.6 Pro helped disprove a 22-year-old graph theory conjecture — basedjensen · 2026-07-24
- Qwen 3.6 35B MoE runs on a Xiaomi 12 Pro with 12GB RAM at 2.4 tok/s — Aromatic_Ad_7557 · 2026-07-24
- OpenAI Codex now supports one project across multiple folders — thesaraharminta · 2026-07-24
- Peter Diamandis says Kimi has dominated open-weights leaderboards for a year — PeterDiamandis · 2026-07-24