Opus 5.5 at Medium Beats GPT-6 Astra Max on GDPval-AA at ~80% Lower Cost
nicolechirps · x · 2026-09-23
A comparison post claims that on the GDPval-AA benchmark, Claude Opus 5.5 at the medium setting outperforms GPT-6 Astra at its highest setting.
Estimated cost per task:
- Opus 5.5: $0.85
- GPT-6 Astra: $4.50
That's a better score at roughly 80% lower cost, per the post's screenshot. Benchmark details and sample sizes are not provided in the tweet itself.
Related event: Opus 5.5 mid-tier beats GPT-6 max at one-fifth the cost(3 posts)→
More from Models
- OpenAI Boosts GPT-6 Prompt Caching, Input Tokens Now Up to 90% Cheaper — OpenAIDevs · 2026-09-23
- Rumor that Opus 5.5 was distilled from a larger teacher model sparks debate — BLUECOW009 · 2026-09-23
- Anthropic's Claude Opus 5.5 system card adopts external evaluation-awareness framework — maksym_andr · 2026-09-23
- GPT-6 Terra Spotted Listed in Hermes — Official Integration or Placeholder? — eugenetel · 2026-09-23
- Claude's Reset Button Now Live on Web and Desktop, Mobile Still Pending — edwinarbus · 2026-09-23
- World #2 Chess GM Hikaru Praises Muse for Voluntarily Flagging Its Own Errors — alexandr_wang · 2026-09-23