Code Arena WebDev puts four open-weight Chinese models near the frontier
floriandotorg · reddit · 2026-08-04
Code Arena’s WebDev leaderboard shows OpenAI’s Sol as the only OpenAI model in the top 10, while four open-weight Chinese models have reached frontier-quality performance.
- The screenshot highlights rankings for models across front-end web development tasks.
- It includes both score and price columns, making the cost/performance contrast visible.
- One of the open-weight Chinese models is described as cheap enough to run for days for the cost of one Opus task.
- The post frames the situation as a major shift in who is competitive on web dev benchmarks.
Related event: Chinese Open-Source Models Challenge Leaders in Code Arena WebDev(2 posts)→
More from Models
- xAI updates Grok Build with Grok 4.5, skills, MCP, and plan mode — elonmusk · 2026-08-04
- RL on custom search harnesses may beat the “one big model” idea — shangbinfeng · 2026-08-04
- Qwen 3.8 Max reaches 42% on the hard INDUCTION benchmark, taking second place — DeryaTR_ · 2026-08-04
- OpenAI hires the creator of WebRTC as its GPT-Live voice system gets a deep dive — bookwormengr · 2026-08-04
- GPT-5.6 Misspells Email Address, Then Hallucinates a Post-Hoc Excuse — WolframRvnwlf · 2026-08-04
- NousResearch ships Hermes Agent v0.20.0 with live voice streaming and wake words — No_Afternoon_4260 · 2026-08-04