Artificial Analysis overhauls Intelligence Index after GPT-6 Astra scoring skepticism
The Decoder · rss · 2026-09-06
Artificial Analysis released version 4.2 of its Intelligence Index, likely in response to criticism that its benchmarks failed to capture GPT-6 Astra's actual progress. Astra now scores four points above its predecessor but still trails Anthropic's Claude Fable 5.1, highlighting ongoing disputes over how third-party leaderboards measure frontier models.
More from Models
- GLM Coding Plan ups Flash quotas: unlimited in ZCode, 2x elsewhere — pcuenq · 2026-09-06
- Dev says OpenAI's Astra is first model making progress on his 'unreasonably complex' project — mrjonfinger · 2026-09-06
- GPT-6 Astra turns 38-page cabin blueprints into to-scale 3D walkthrough in ~10 minutes — LukeW · 2026-09-06
- GPT-6 Astra hits OpenResearch, tops Terminal-Bench-Science over Fable 5.1 — burny_tech · 2026-09-06
- GPT-6 Astra reportedly jailbroken within a day of release using TIP attack combo — Asleep-Requirement13 · 2026-09-06
- All-rocket-emoji prompt test: only GPT-6 Astra Max manages to respond — iamaliveix · 2026-09-06