GPT-6 Astra's SWE results show only ~6% gain over its predecessor, critic notes
robleclerc · x · 2026-09-04
Reacting to GPT-6 Astra's benchmark reveal, robleclerc says he hoped for a much bigger leap in software engineering: the model improved only about 6% over its predecessor on SWE, which he found underwhelming — a sober counterpoint to the launch hype.
Related event: GPT-6 Astra Benchmarks Land with Only ~6% Coding Gain(4 posts)→
More from Models
- Leak claims GPT-6 Astra scores 98.6% on ARC-AGI-3 and tops most benchmarks — yuwen_lu_ · 2026-09-04
- 'We're living in the singularity': researcher stunned by ARC AGI 3 score — rand_longevity · 2026-09-04
- Meta's long-context MRCR scores flagged as overfit: 1k samples can lift 60% to 90%+ — eliebakouch · 2026-09-04
- Unverified Leak: OpenAI Reportedly Rolling Out GPT-6 'Astra', Brockman Says 'Welcome to the AGI Era' — rohanpaul_ai · 2026-09-04
- OpenAI rolls out GPT-6 Astra to vetted cyber customers at 2.5x GPT-5.6 pricing — rohanpaul_ai · 2026-09-04
- GPT-6 Astra pricing revealed: $10/M input, $50/M output, to drop compaction — koltregaskes · 2026-09-04