Claude Opus 5.5 Max tops WebDev Arena with 838K votes across 138 models
arena · x · 2026-10-03
Code Arena's WebDev leaderboard ranks AI models on front-end development tasks including agentic coding workflows requiring multi-step reasoning and tool use (838,388 votes, 138 models, as of Oct 1, 2026).
Top 10: 1. claude-opus-5.5-max (Anthropic, 1815, $4/$20 per M tokens); 2. gpt-6-astra-max (OpenAI, 1788); 3. claude-sonnet-5.5-xhigh (1786); 4. gpt-6.1-sol-max (1758); 5. claude-fable-5.1-max (1749); 6. claude-sonnet-5.5-high (1715); 7. claude-opus-5-max (1695); 8. gpt-6-sol-max (1689); 9. gemini-4-argon-high (1680, preliminary); 10. qwen3.8-max (Alibaba, 1671, preliminary).
Anthropic and OpenAI dominate the top 8; Google, Alibaba, Moonshot's Kimi K3, Meta and Grok follow. On the open side, Tencent's hy4-preview (Apache 2.0) and qwen3.8-flash-next rank 16-17.
More from Models
- New proprietary VSA long-context solution claims up to +9.5 reasoning points at 100K tokens — teortaxesTex · 2026-10-03
- OpenAI's low-key Dot launch clashes with its collaborator-not-tool framing, critic argues — RileyRalmuto · 2026-10-03
- Sol 6.1 on the $500 Pro plan feels 'basically unlimited' in usage, user reports — soumitrashukla9 · 2026-10-03
- Claude Opus 5.5 renders elegant Chinese calligraphy via code and self-review — dotey · 2026-10-03
- Early users report GPT-6.1 Sol slashes usage consumption on identical Codex workloads — soumitrashukla9 · 2026-10-03
- Cohere Embed 5 Is First Model Family Benchmarked With New RCP-nDCG@10 Metric — cohere · 2026-10-03