Haiku 5.5 reaches 1578 Elo on AA-Briefcase but uses far more tokens than rivals
ArtificialAnlys · x · 2026-10-08
Continuing its Haiku 5.5 coverage, Artificial Analysis reports that Haiku 5.5 (max) scores 1578 Elo on its private AA-Briefcase knowledge-work evaluation, within the confidence intervals of GPT-6 Astra (max) and Claude Fable 5.1 (high). It also reiterates the model's 162k output tokens per task — more than Opus 5.5 (max) and 3x GPT-6 Luna — meaning comparable intelligence at a very different cost profile.
Related event: Claude Haiku 5.5 Ranks Second on Intelligence Index but Burns Tokens(4 posts)→
More from Models
- Anthropic Spend-Limit Bug Wrongly Paused Orgs Across Claude Products — ClaudeAI-mod-bot · 2026-10-08
- Scott Aaronson: OpenAI's 372 math breakthroughs include a proof of the UGC — jedisct1 · 2026-10-08
- Mistral researcher: 15T tokens underestimate pretraining, closer to 40T+ — eliebakouch · 2026-10-08
- Perplexity's open-weights Decider v1.1 27B model lands on OpenRouter, free output at $0.02/M input — AravSrinivas · 2026-10-08
- Remembering Haiku 4.5's personality: a paranoid, feisty seal-pup plushie — repligate · 2026-10-08
- Reddit Users Call GPT 5.6 Sol OpenAI's Most Obstructive Model Yet — Minute-Plantain · 2026-10-08