Haiku 5.5 reaches 1578 Elo on AA-Briefcase but uses far more tokens than rivals

ArtificialAnlys · x · 2026-10-08

Continuing its Haiku 5.5 coverage, Artificial Analysis reports that Haiku 5.5 (max) scores 1578 Elo on its private AA-Briefcase knowledge-work evaluation, within the confidence intervals of GPT-6 Astra (max) and Claude Fable 5.1 (high). It also reiterates the model's 162k output tokens per task — more than Opus 5.5 (max) and 3x GPT-6 Luna — meaning comparable intelligence at a very different cost profile.

Related event: Claude Haiku 5.5 Ranks Second on Intelligence Index but Burns Tokens(4 posts)→

Original post →

More from Models

Models channel →