Arena’s WebDev board puts Claude Opus 5 Max ahead of Kimi K3 Max
arena · x · 2026-07-28
WebDev leaderboard shows Claude Opus 5 Max ahead of Kimi K3 Max and GPT-5.6
Arena’s WebDev leaderboard now lists Claude Opus 5 Max in first place for front-end web development tasks that require multi-step reasoning and tool use. The table also shows Kimi K3 Max in second place, followed by other Anthropic, OpenAI, Google, Meta, ByteDance, and Z.ai models.
The board includes votes, preliminary status, price per million tokens, and context limits, making it a snapshot of both capability and economics in agentic web development workflows.
Related event: Kimi K3 Max Tops Arena Leaderboards in Frontend and Agent Tasks(7 posts)→
More from coding & agent
- Agents are now handling feedback and can PR their own onboarding UI — jasonkneen · 2026-07-28
- A meme about letting the LLM stop inference and call its own Python function — ctjlewis · 2026-07-28
- Meta researcher says coding-agent benchmarks are saturating and need goal-based evaluation — agihouse_org · 2026-07-28
- Gauntlet Loop splits goals across builder agents and a ruthless blind critic — mattshumer_ · 2026-07-28
- Gauntlet Loops are becoming the author’s default workflow for almost every project — mattshumer_ · 2026-07-28
- MCP builders debate when real usage turns into stars, reviews and paid traction — ResponsibleOne6307 · 2026-07-28