Claude Opus 5 peaks at medium effort on FrontierCode, while max can hurt coding quality

量子位 · wechat · 2026-07-27

Opus 5’s best coding performance shows up at medium, not max

A test on the FrontierCode benchmark swept Claude Opus 5 across inference-effort levels from low to max and found a counterintuitive curve: medium delivered the best performance, while cranking effort all the way up did not help and could even hurt.

The recommended strategy is to pick one effort level per workflow and keep it stable to preserve cache hits. For many structured tasks, the best tradeoff appears to be medium, not max.

Original post →

More from coding & agent

coding & agent channel →