GPT-6 Astra Tops Code Arena WebDev at 1,797, Qwen 3.8 Flash Next Cracks Top 10
pbaylies · x · 2026-09-06
Code Arena's latest WebDev leaderboard shows GPT-6 Astra (Max) at #1 with 1,797 points, 35 points above Claude Fable 5.1 (Max). The benchmark has models plan with tools and build live web apps, with users comparing paired outputs. Notably, Astra leads OpenAI's previous entry GPT-5.6 Sol (xHigh) by 180 points, which now sits at #13. Commenters were also surprised to see Qwen 3.8 Flash Next in the top 10 alongside Fable and Grok.
Related event: GPT-6 Astra Tops Code Arena WebDev Leaderboard(3 posts)→
More from Models
- Hands-on GPT 6 Astra review: real work, no game demos — Rasmic · 2026-09-06
- Astra day-two impressions: chattier, strong spatial sense, less jargon than Fable 5 — bindureddy · 2026-09-06
- 200 tok/s on 8GB VRAM: dev benchmarks 6 small models for local AI — TheMoonMidas · 2026-09-06
- Humanize + GPT-5.5 solves 670/672 Lean-verified proofs, tops PutnamBench at 99.7% — songhan_mit · 2026-09-06
- "Claude Solved Navier–Stokes" Is Just a Rumor: No Paper, No CMI Submission, Say Fact-Checkers — johnseach · 2026-09-06
- Frontier AI models begin crossing the human baseline on SimpleBench — Bojackin_Around · 2026-09-06