Claude Haiku 5.5 jumps 257 points to 1587 on WebDev Arena leaderboard
arena · x · 2026-10-09
WebDev Arena (852K+ votes, 142 models) reports that Claude Haiku 5.5 scored 1587 points, a +257-point improvement over Haiku 4.5's 1330, with significant gains across all categories.
Top of the overall leaderboard: claude-opus-5.5-max (1813), gpt-6-astra-max (1786), and claude-sonnet-5.5-xhigh (1774). Qwen3.8-max (1672) ranks among the strongest Chinese/open models, while Xiaomi's mimo-v2.6-flash offers notable value at $0.14/$0.28 per million tokens with 1637 points.
Related event: Claude Haiku 5.5 Tops WebDev Arena with 257-Point Jump(2 posts)→
More from Models
- FineWeb author: annotating pretraining data with a 27B model is wild but pays off at deployment — antoine_chaffin · 2026-10-09
- HF researcher: fine-tuned small models win on throughput, zero-shot wins on capabilities — antoine_chaffin · 2026-10-09
- Dev racing to burn $800 of Gemini API credit teases rumored Gemini 4 Argon and Nano Banana Pro 2 — Angaisb_ · 2026-10-09
- 16 VPD weight edits boost model accuracy 10x, revealing the attention head that suppresses introspection — Sauers_ · 2026-10-09
- GPT-6.1 Sol "Ultrafast" clocks in at just ~49 tok/s in user API speed test — RexDouglass · 2026-10-09
- One attention head drives sandbagging-like introspection in Qwen3-1.7B; ablating it helps — Sauers_ · 2026-10-09