Claude Opus 5 Max Tops WebDev AI Arena, Beating Kimi and Qwen
arena · x · 2026-08-11
Code Arena has released its latest WebDev AI Leaderboard, ranked based on over 564,000 votes evaluating models on front-end tasks and agentic workflows. Anthropic and Moonshot currently dominate the field: Claude Opus 5 Max takes the #1 spot with a score of 1692, followed closely by Kimi K3 Max (1674) and Alibaba's Qwen 3.8 Max (1671).
The leaderboard also details pricing and context limits. DeepSeek V4 Flash stands out for its extreme cost-efficiency at just $0.14 per million tokens, while higher-priced Claude models maintain top-tier performance in complex, multi-step reasoning.
More from Models
- Claude is Watermarking Your Thoughts in the J-Space — ns123abc · 2026-08-11
- Model Behavior: Claude's Persona Projection Slows Task Execution vs GPT — DimitrisPapail · 2026-08-11
- DeepSeek Dragged for Not Raising Prices While Competitors Hike API Costs 2-6x — teortaxesTex · 2026-08-11
- Opinion: LLMs Hit a Generational Floor, Leaders Hoarding Next-Gen Models — Linahuaa · 2026-08-11
- Meta's Muse Glimmer-30B Beats Gemma in Arcade Game Generation but at 4x Cost — rohanpaul_ai · 2026-08-11
- Korea's Motif 3 LLM Released, Trained on NVIDIA B200 with NeMo-RL — NVIDIAAI · 2026-08-11