After a full day testing Opus 5, the author sees little improvement

jiayuan_jy · x · 2026-07-26

The author says they spent a full day testing Opus 5 and did not notice any significant improvement.

Their takeaway is that it mostly just gets the job done, with performance feeling similar to GPT 5.6 sol and Opus 4.8.

Related event: Opus 5 Shows No Significant Improvement, Only Lower Failure Rate(2 posts)→

Original post →

More from Models

Models channel →