Qwen3.8-Max Tops Code Arena, Beats Claude in Value
airesearch12 · x · 2026-09-02
Alibaba's Qwen3.8-Max-0902 debuted at #1 overall in Code Arena: WebDev with 1,691 pts, becoming the highest-scoring model on the Pareto frontier at a blended cost of $5/MToken. With this release, three previous top models—Claude Opus 5, Kimi K3, and Qwen3.8—were moved off the Pareto frontier despite remaining #2–#4 overall. The update highlights significant progress in code generation capabilities and pricing advantages for Qwen.
More from Models
- Polymarket puts 78% odds on OpenAI's rumored Astra model launching tomorrow — Polymarket · 2026-09-03
- Polymarket puts 78% odds on OpenAI's rumored Astra model shipping tomorrow — Polymarket · 2026-09-03
- 'GPT-6-ASTRA' spotted staged on the OpenAI API, unconfirmed — ThunderBeanage · 2026-09-03
- Google DeepMind releases Gemini 3.8 Flash and 3.8 Flash Cyber — Google DeepMind · 2026-09-03
- FrontierHarness eval: 12 agent harnesses, same model — pass rates span 50%–67%, cost per pass $1.05–$18.34 — stuffyokodraws · 2026-09-03
- Gemini 3.8 Flash praised as 'another good model' with improving science capabilities — vivnat · 2026-09-03