Unverified: GPT-6 Astra reportedly scores 91.8% on SpatialBench vs 80% human baseline
VraserX · x · 2026-09-06
A blogger claims GPT-6 Astra scored 91.8% on SpatialBench, beating an 80% human baseline on spatial reasoning — described as one of the last areas where humans held a clear edge. The claim comes from a third-party account and is unverified by OpenAI; treat the benchmark figure as rumor.
More from Models
- Providers wage price war over serving DeepSeek-v4-flash, user burns massive tokens for a few dollars — MaziyarPanahi · 2026-09-06
- Annoyed by 'maximum length' prompts? Refreshing the page lets you keep chatting — hi-sci-collab · 2026-09-06
- GPT-6 builds a full 3D Roguelike level in Godot using just 3% of a 20X quota — op7418 · 2026-09-06
- Dev shows Astra's unusual habit of verifying its own work, unlike any other model — dkundel · 2026-09-06
- Codex Business users can't buy extra resets as weekly limits hit 60% — cnakazawa · 2026-09-06
- DeepSeek v4 models reportedly 30% off via third-party channel, unverified — matlabulous · 2026-09-06