图表显示 Claude Opus 5 在协议任务上领先同列版本

nlarusstone · x · 2026-07-25

配图是一张 Claude 各版本在协议/任务上的对比图。

图中 Claude Opus 5 在标注为 Understanding (Benchling) 的场景里拿到最高值 0.78,高于 Claude Sonnet 5 的 0.60 以及图中其他列出的 Claude 变体。

从整体柱状图看,Opus 5 相比早期版本有提升,但这条帖子本身没有给出完整的 benchmark 背景。

所属事件:Anthropic 发布 Claude Opus 5,多项基准领先(75 条相关)→

原文链接 →

「模型」频道最新

更多「模型」频道 AI 资讯 →