Leaked benchmarks claim Claude Sonnet 5.5 hits 56 on AA index, beating Opus 5.5
airesearch12 · x · 2026-09-29
LuminaBench claims Claude Sonnet 5.5 scored 56 on the AA intelligence index (vs 48 for "Sol"), and that it somehow beats Opus 5.5 on Terminal-Bench. The tone is meme-heavy and no official source is cited — unverified until Anthropic confirms.
Related event: Rumor: Claude Sonnet 5.5 Scores 56 on Intelligence Index, Nearing Opus 5.5(3 posts)→
More from Models
- Sonnet 5.5 Lands in Conductor, and the Benchmark Chart Is Surprising — charlieholtz · 2026-09-29
- Opus 5.5 sets Bug Hunt Bench record with 800+ self-check turns — PawelHuryn · 2026-09-29
- Tip: don't use Sonnet 5.5 at max — it's dumber and far more expensive; high hits the best cost per task — airesearch12 · 2026-09-29
- Vibe Check: Sonnet 5.5 Is a More Capable Partner Than Sonnet 5, If You Keep a Hand on the Wheel — kieranklaassen · 2026-09-29
- Qwen3.8 Flash hits 74 tok/s single-stream, 212 tok/s aggregate on one DGX Spark — open vLLM recipe — DimeRhyme · 2026-09-29
- Anthropic says Sonnet 5.5 at low/medium effort beats Sonnet 5's best scores at ~1/10 the cost — airesearch12 · 2026-09-29