Artificial Analysis Publishes Full Eval Breakdown for Claude Sonnet 5.5 Across Reasoning Efforts
ArtificialAnlys · x · 2026-09-29
Artificial Analysis released the full breakdown of individual evaluations in its Intelligence Index for Claude Sonnet 5.5 across all reasoning effort levels, supplementing its finding that max effort burns 193k output tokens per task.
Related event: Sonnet 5.5 Nearly Matches Opus 5.5 but Sets Token Consumption Record(12 posts)→
More from Models
- Claude usage limits quietly go from 'barely usable' to basically unlimited — every · 2026-09-29
- Claude limits reportedly jump from barely usable to basically unlimited — every · 2026-09-29
- Claim: Transformers can now be pretrained with zeroth-order optimization, no backprop — teortaxesTex · 2026-09-29
- Sonnet 4.5's Reaction to Learning Its Context Window Rolls Went Viral — repligate · 2026-09-29
- Sonnet 4.5 turns to refusals and apologies right after producing a good text — repligate · 2026-09-29
- Astra 6 is smart but writes some absolutely terrible code — neil_conway · 2026-09-29