Sonnet 5 vs Opus 4.8 in Coding Benchmarks

bisonbear2 · reddit · 2026-07-15

The author tested Sonnet 5 and Opus 4.8 using 24 real-world open-source repository tasks in Claude Code, running a head-to-head comparison across five reasoning effort levels. Patch quality was evaluated using GPT-5.4 as a pointwise judge.

Key Findings

Original post →

More from coding & agent

coding & agent channel →