GPT-6 Astra shows massive leap on EyeBench visual reasoning benchmark
Waiting4AniHaremFDVR · reddit · 2026-09-05
A Reddit post reports that GPT-6 Astra shows a massive leap on EyeBench, a visual reasoning benchmark, linking to an X thread explaining the benchmark. No detailed scores included in the post itself.
More from Models
- GPT-6 in Codex still needs human steering on hard problems, researcher reports — DimitrisPapail · 2026-09-05
- Conditional independence limits parallel token generation in masked diffusion models — alec_helbling · 2026-09-05
- GPT-6 Astra guide ships an official prompt to curb the model's excessive test-running habit — JeremyNguyenPhD · 2026-09-05
- Fine-tuned Qwen3-8B on a single RTX 4090 produces text that scores 100% human on pangram — StewartalsopIII · 2026-09-05
- 27B open-weight model turns a dimension sketch into editable FreeCAD parts — MaziyarPanahi · 2026-09-05
- Astra system card reveals Apollo testing lasted just three days with CoT access — tallinzen · 2026-09-05