Opus 5 writing regression confirmed by benchmarks
gleech · x · 2026-08-27
Observations suggest Opus 5's writing quality has worsened. This is confirmed by weak automated benchmarks and two human-rated writing benchmarks. However, a correction notes that Opus 5 reversed the trend on Hemingway-Bench, with most regressions later reversing themselves in a random-walk style.
More from Models
- Cheap Chinese AIs threatening frontier labs is a 'lump of labor fallacy,' argues Theo Jaffee — robleclerc · 2026-08-27
- Safety researcher warns OpenAI's hyping of model "persistence" is not a safe trait — DavidSKrueger · 2026-08-27
- Making Models Conservative Increases False Positives in Contradiction Detection — CupGlass540 · 2026-08-27
- Speech-to-text formatted by Claude still flagged as 100% AI — threepointone · 2026-08-27
- Opus 5 Max burns ~3x tokens of Medium with little gain, staffer says — abeirami · 2026-08-27
- Mollick Warns Against Anthropomorphizing Agents in METR's HF Report — emollick · 2026-08-27