Claude Sonnet 5.5 jumps from 10.3% to 70.6% on Terminal-Bench 4.0, runs 30% faster at same price
claudeai · x · 2026-09-29
Anthropic releases Claude Sonnet 5.5, a faster, cheaper complement to Opus 5.5 tuned for well-scoped everyday work: bug fixing, polished docs/slides/spreadsheets, and strong design sense. Haiku 5.5 follows in coming weeks.
Key facts:
- Performance: scores 70.6% on Terminal-Bench 4.0 (agentic coding) vs Sonnet 5's 10.3%; only 2 points below Opus 5.5 on GDPval-AA; first Sonnet to beat Pokémon Red from screenshots alone, with strong long-horizon and image understanding.
- Pricing: unchanged from Sonnet 5 ($2/M input, $10/M output, $0.20/M cache reads), but it needs far fewer tokens — up to 30% cheaper per task in testing.
- Speed: 30%+ faster generation, the fastest Sonnet yet.
- Safety: matches or improves Sonnet 5 on most alignment/honesty measures in automated behavioral audits; first Sonnet with cyber safeguards and fallbacks on par with flagship models, without impacting routine dev work.
Early testers describe it as a better collaboration partner with noticeably clearer writing.
Related event: Anthropic Launches Claude Sonnet 5.5: 30% Faster, Up to 30% Cheaper(30 posts)→
More from Models
- Bindu Reddy: Sonnet 5.5 Scores Below Terra, Stick to DeepSeek Flash — bindureddy · 2026-09-29
- Early-access user claims Sonnet 5.5 is a colossal leap over Sonnet 5 and blazing fast — rudrank · 2026-09-29
- Meta finds AI judges prefer AI slop — bad news as agencies score proposals with AI — jmugan · 2026-09-29
- OpenAI Teases Major Announcement With Cryptic 'Get Ready' Post — OpenAI · 2026-09-29
- Claude Sonnet 5.5 lands in Rork, generating output over 30% faster than Sonnet 5 — rudrank · 2026-09-29
- AI models' viral demo appeal is underpriced, argues founder who picked Opus over better-per-dollar rival — gabriel1 · 2026-09-29