Claude Opus 5 曝光:多榜单综合登顶成新 SOTA
Claude Opus 5 在最新曝光的多个第三方基准测试榜单中表现出色,综合能力登顶成为新的全局 SOTA(当前最优模型)。在 Artificial Analysis 和 BenchmarkList 的评测中,它均超越了 Claude Fable 5 等竞品,引发了社区的广泛关注。
已确认
根据 Artificial Analysis 更新的榜单,Claude Opus 5 在 Intelligence Index 上获得了 61 分,综合排名第一,略高于 Claude Fable 5。此外,@davidthesong 提供的 BenchmarkList 截图显示,Claude Opus 5 被标记为新的全局 SOTA 第 1 名,其覆盖了 52 个基准,实验性 ECI 评分为 154.80。在衡量现实工作表现的 GDPval-AA v2 榜单中,同样有截图证实 Opus 5 的排名领先于 Fable 5。
尚未确认
@Hesamation 指出,尽管 Claude Opus 5 在综合智能指数上登顶,但在编程代理(Coding Agent)细分能力上,Fable 5 依然保持领先。
为什么重要
Claude Opus 5 在综合智能指数上的登顶,标志着大模型能力天花板的再次提升。同时,Opus 5 与 Fable 5 在通用能力与编程代理任务上的差异化表现,为开发者在不同应用场景下的模型选择提供了明确的参考依据。
2026-07-25 ~ 2026-07-25 · 6 条相关
- 第 1 集:Claude Opus 5 曝光:多榜单综合登顶成新 SOTA(2026-07-25,6 条)
- 第 2 集:Opus 5 与 GPT-5.6 Sol 实测:能力趋同,成本与风格成选型关键(2026-07-25,7 条)
- 第 3 集:Claude Opus 5 视觉评测落后且成本高昂(2026-07-25,4 条)
一手来源
- Claude Opus 5 综合登顶,Fable 5 仍领跑编程代理 — Hesamation ·
- Claude Opus 5 登顶 BenchmarkList 成为新全球 SOTA — davidthesong ·
- Artificial Analysis 榜单显示 Claude Opus 5 领先 Fable 5 — Leonardo-editing · 2026-07-25
- Claude Opus 5 在模型能力图上位居前列 — scaling01 · 2026-07-25
- 【源头】Claude Opus 5 综合登顶,Fable 5 仍领跑编程代理 — Hesamation · 2026-07-25
- 【源头】Claude Opus 5 登顶 BenchmarkList 成为新全球 SOTA — davidthesong · 2026-07-25
- Claude Opus 5 在 Artificial Analysis 上略胜 Claude Fable 5 — thesaraharminta · 2026-07-25
- Claude Opus 5 登顶 Artificial Analysis,分数达到 61 — Rare_Bunch4348 · 2026-07-25