Researcher claims frontier LLM coding has plateaued since Opus 4.8, still rates Astra higher
Yuchenj_UW · x · 2026-09-27
Yuchen Jin pushes back on the hype around Opus 5.5, saying he still doesn't think it beats Astra. His bolder claim: frontier LLM coding capability has plateaued since Opus 4.8, with no meaningful jump since.
More from Models
- LeCun: LLMs are mostly information retrieval systems, not thinkers — ylecun · 2026-09-27
- Persona vectors emerge at 0.22% of pretraining and persist into post-trained models, NeurIPS paper shows — burny_tech · 2026-09-27
- Claude 3 Opus Finds Zero-Days in Source Code, Sparking AI Risk Debate — JasonDClinton · 2026-09-27
- Tokens Keep Getting Cheaper Per Usefulness — Your View of What's Possible Is Stale — avt_im · 2026-09-27
- Jev takes 27% of OpenRouter classification requests, 2x DeepSeek V4 Flash — gaganghotra_ · 2026-09-27
- Claude Opus 5.5 makes its own 20-page sketchbook: handwriting, doodles, and piano music — CurieuxExplorer · 2026-09-27