AA 实测:Claude Sonnet 5.5 单任务烧 19.3 万输出 token,是 GPT-6 Astra 的 7 倍

ArtificialAnlys · x · 2026-09-29

Artificial Analysis 测量发现,Claude Sonnet 5.5(max effort)在每道 Intelligence Index 任务上平均消耗约 19.3 万输出 token,是所有已测模型中最高的,约为 GPT-6 Astra(max)的 7 倍——用超大推理量换来了顶尖表现。

但在更低 effort 档位上,性价比劣势明显:low/medium/high 三档在「智能 vs 单任务输出 token」的权衡曲线上,分别落后于 GPT-6 Sol 的 high、xhigh、max 档,即同档推理预算下 Sol 能用更少 token 达到更高性能。

所属事件:AA 评测 Sonnet 5.5:性能逼近旗舰,token 消耗最高(8 条相关)→

原文链接 →

「模型」频道最新

更多「模型」频道 AI 资讯 →