Opus 5 发布,agentic 编码、搜索和电脑操作成绩强势
TheInfiniteUniverse_ · reddit · 2026-07-25
帖子宣布 Opus 5 刚刚发布,配图是一张完整的基准测试对比表,覆盖:
- agentic terminal coding
- knowledge work
- agentic search
- multidisciplinary reasoning
- computer use
- agentic coding
- business workflows
- legal / health / biology
从图里看,Opus 5 在多项基准上都拿到领先或接近领先的成绩,尤其是在 agentic search、computer use、知识工作和若干专业场景上表现突出;表中同时对比了 Fable 5、Opus 4.8 和 GPT-5.6 Sol。
所属事件:Anthropic 发布 Claude Opus 5:性能达 SOTA 且价格减半(106 条相关)→
「模型」频道最新
- 用户称 ChatGPT Image 2 在修图和准确性上已反超 Gemini — dreamwieber · 2026-07-25
- Grok 4.5 在 Augment 里录得最大周增幅 — Daniel_Farinax · 2026-07-25
- 有人能让 AI 以 24fps 看完整段视频吗 — mattshumer_ · 2026-07-25
- Kimi 权重若公开,争论会转向硬件经济学对 V4 — teortaxesTex · 2026-07-25
- Anthropic 详解 Fable 5 编排、advisor 模式与缓存成本 — brada · 2026-07-25
- PaddlePaddle 的 HPD-Parsing 在 Hugging Face 走热,主打文档解析 — PaddlePaddle · 2026-07-25