MolmoSpaces benchmark launched: GPT-Astra beats all open-source VLA baselines zero-shot
notmahi · x · 2026-09-15
The Molmo team released MolmoSpaces-v1, a benchmark for robotic spatial understanding, asking whether robotics is having its "GPT moment."
Key result: GPT-Astra outperformed all open-source VLA/world-model baselines on a zero-shot subset of the benchmark. The authors published execution traces and argue general-purpose models are showing a leap in spatial reasoning for robotics.
More from Embodied
- Physical Intelligence unveils OM-1, a robot foundation model trained purely on human data — zipengfu · 2026-09-15
- Eren Chen launches robot hardware shop: Agibot humanoids from $29,999, US shipping — chris_j_paxton · 2026-09-15
- OpenAI acquires camera startup Glass Imaging for over $300M, WSJ reports — rohanpaul_ai · 2026-09-15
- Amid new robotics launches, a pointer to best practices for policy evaluations — eigenron · 2026-09-15
- "Learning robotics today is like learning to code in 2010" — Paimaamu · 2026-09-15
- Mecha Corp unveils Spike, an RL-trained bipedal robot running purely on proprioception — Scobleizer · 2026-09-15