Arena’s WebDev board puts Claude Opus 5 Max ahead of Kimi K3 Max
arena · x · 2026-07-28
WebDev leaderboard shows Claude Opus 5 Max ahead of Kimi K3 Max and GPT-5.6
Arena’s WebDev leaderboard now lists Claude Opus 5 Max in first place for front-end web development tasks that require multi-step reasoning and tool use. The table also shows Kimi K3 Max in second place, followed by other Anthropic, OpenAI, Google, Meta, ByteDance, and Z.ai models.
The board includes votes, preliminary status, price per million tokens, and context limits, making it a snapshot of both capability and economics in agentic web development workflows.
More from coding & agent
- Looking for a classifier of software engineering task shapes to pick models per task — StewartalsopIII · 2026-09-11
- Steal this idea: prompt-to-hardware where agents assemble custom devices — paraschopra · 2026-09-11
- Model Is the Least Interesting Part: A Guide to Six Core AI Architectures from RAG to Multi-Agent — goyalshaliniuk · 2026-09-11
- Non-coder builds layered memory architecture: 20k tokens tracks a year of agent conversations — matteoianni · 2026-09-11
- Warp's six non-engineering teams all run on Linear and Claude Code — mon__lim · 2026-09-11
- 9-year backend dev: AI code isn't the problem, the rate of making a mess is — Sweaty-Landscape-561 · 2026-09-11