Thread ranks frontier coding AI: Opus 4.5 at average SWE, next-gen above experts

menhguin · x · 2026-09-07

In a discussion on how capable frontier models have become, @menhguin offers a tiered take: Claude 3.5 Sonnet already clears the "above the median person" bar, Opus 4.5 roughly clears the "average SWE" bar, and rumored next-gen models ("Astra" or "Fable") may reach "likely better than an expert in a sensitive technical operation" level.

Original post →

More from Models

Models channel →