Yuchen Jin: Opus 5.5 underwhelms, frontier LLM coding has plateaued
Yuchenj_UW · x · 2026-09-23
AI researcher Yuchen Jin shared first-day impressions of Opus 5.5 and came away unimpressed: it got a tricky research question wrong that Astra answered correctly, and Astra still feels better for coding.
He doubles down on his earlier take that frontier LLM coding capability has plateaued, arguing the real battle is now about intelligence per dollar rather than raw capability.
More from Models
- Grok 4.7 flops in 100 multi-agent coding evals despite insightful solutions — teortaxesTex · 2026-09-23
- Will rumored GPT-6 'Sol' actually ship inside ChatGPT? — flowersslop · 2026-09-23
- Claude Opus 5.5 Said to Fall Back on Frontier LLM Dev Tasks, Drawing Fire — basedjensen · 2026-09-23
- Higgsfield demo: Claude Opus 5.5 crushes GPT-6 Astra at 3D game generation — VraserX · 2026-09-23
- User Gives Opus 5.5 Creative Tools and Asks What It Dreams About — angrypenguinPNG · 2026-09-23
- Forward Future puts Opus 5.5 through 8 tests: cities, games, animation — MatthewBerman · 2026-09-23