Flagship models put to the test: rebuilding Geisel Library in 3D from photos alone
Lianhuiq · x · 2026-09-04
EnactraAI benchmarked frontier flagships (Fable 5.1, Gemini 3.8 Flash, Muse Spark 1.3) on a hard spatial-simulation task: reconstructing Geisel Library in 3D from photos only. The models differed markedly in geometry, structure, and visual fidelity, with a slot reserved for the much-anticipated GPT-6 Astra.
More from Models
- Screenshots of Amazon Astra's Chain-of-Thought Output Circulate on Reddit — Tough_North7059 · 2026-09-04
- Theo: Astra is world-class at everything except frontends and mergeable code — adonis_singh · 2026-09-04
- OpenAI says GPT-6 Astra shows Critical cyber capabilities, forcing harder ExploitBench evals — SIGKITTEN · 2026-09-04
- eyebench-v3 adds Muse-Spark 1.3 at 33%, Gemini-3.8-Flash flat vs predecessor — adonis_singh · 2026-09-04
- Zvi warns Astra's CoT controllability surge could systematically erode AI monitorability — TheZvi · 2026-09-04
- GPT-6 Astra barely improves over 5.6 Sol on Artificial Analysis Intelligence Index — burny_tech · 2026-09-04