GPT-6 Astra trails Fable by 10 points and Sol by 7 on Agentic Index
Capable-S · reddit · 2026-09-04
A Reddit user points out that despite GPT-6 Astra's impressive traditional benchmark scores, it trails Fable by 10 points and GPT 5.6 Sol by 7 on the Agentic Index, performing at roughly the level of Terra and Qwen3.8 27B on agentic tasks — a stark contrast to the launch hype.
More from Models
- OpenAI's Astra can now layout and route PCBs, sparking hardware engineering debate — MikePFrank · 2026-09-04
- Ethan Mollick: Astra just takes action, spinning up agents on vague requests — emollick · 2026-09-04
- 'AGI is 74% deepswe': GPT-6-Astra benchmark results become an AI-circle meme — amaarora · 2026-09-04
- Grok 4.7 reportedly days away, trained on SpaceX engineering data; Grok 4.6 already ties GPT-6 Astra at 61 — XFreeze · 2026-09-04
- Ex-OpenAI safety lead Miles Brundage: if your primary emotion on AI isn't concern, you're misreading it — Miles_Brundage · 2026-09-04
- Gary Marcus on GPT-6 Astra: symbolic world models are vindication, but no proof of AGI — GaryMarcus · 2026-09-04