Kimi K3 在 ExploitBench 仅得 32%,远落后美国前沿模型

The Decoder · rss · 2026-07-24

The Decoder 引述英国 AI Security Institute 和美国 AI Standards and Innovation Center 的测试称,Moonshot AI 的 Kimi K3 在进攻性网络安全任务上明显落后于美国前沿模型。

关键信息:

文章还指出,Kimi K3 在通用基准上表现不错,但在 cyber 能力上掉队,这与外界关于 Moonshot 可能蒸馏 Anthropic 模型的指控形成了呼应。

原文链接 →

「模型」频道最新

更多「模型」频道 AI 资讯 →