Anthropic pushes back on Sonnet 6 regression claims, says it's a flatly better model
pvncher · x · 2026-09-23
Responding to widespread claims that Sonnet 6 regressed vs 5.6, Anthropic team member pvncher says the new model should be "a flatly better model despite what the evals say," and urges users to report behavior issues via /feedback, noting the team is monitoring submissions closely and responding nimbly.
More from Models
- Matt Shumer declares 'Anthropic has won,' calls new model incredible — mattshumer_ · 2026-09-24
- Testing the Jeb chatbot: inconsistently biased, not neutral — calibrate it like any classifier — PawarBI · 2026-09-24
- Pokemon benchmark Paradigm 3: Astra generalizes to scrambled maps and fan-made games while rivals memorize — gleech · 2026-09-24
- AI Completes Fan-Made Pokemon Brown in 10K Steps: Real Generalization or Whack-a-Mole? — gleech · 2026-09-24
- Next-gen model names surface: Opus 5.5, Fable 5.1, GPT-6 Astra — labs said to be ~2 months ahead internally — haider1 · 2026-09-24
- AI detector debate: economist argues Pangram is the only reliable tool, cites 0 FPR finding — paulnovosad · 2026-09-24