Sonnet 5.5 benchmarks near Opus 5.5 unevenly, while OpenAI keeps token-efficiency edge

cedric_chee · x · 2026-09-29

Early feedback on Sonnet 5.5 is strong: users praise its speed, clear writing, and iteration feel. Benchmarks are hard to interpret—it's close to Opus 5.5 and even wins some evals, but Opus still owns complex, open-ended work.

Two caveats from the author:

Related event: Sonnet 5.5 Nearly Matches Opus 5.5 but Sets Token Consumption Record(12 posts)→

Original post →

More from Models

Models channel →