Closed-loop spatial understanding from just two wrist cameras, no gripper feedback
ChongZzZhang · x · 2026-09-21
A 4x/12x-speed demo shows a robot achieving closed-loop spatial operation using only plain visual understanding from two cameras on the same wrist — no base/world frame observation, no gripper feedback.
Author's takeaways: closed-loop spatial understanding is off-the-shelf, and reasoning and actionability can be coupled without symbolising.
Related event: GPT6 as Policy: Wrist-Mounted Dual Cameras Drive Zero-Shot Robot Stacking(5 posts)→
More from Embodied
- Human vs Robot Fight at San Francisco's REK Event Hit 'Completely Different', Says Attendee — BLUECOW009 · 2026-09-21
- DIY robot U-BOT's maiden voyage: stuck on a tiny stick, grass proves tougher than expected — _Stocko_ · 2026-09-21
- First-ever Human vs Terminator robot fight pits Frankie LaPenna against a bot — BLUECOW009 · 2026-09-21
- Astra shows any 3D/4D prior can be distilled into VLMs, a new embodied AI paradigm — mariyaivasileva · 2026-09-21
- Legless autonomous food robot sparks debate: fixed-purpose machines land before humanoid cooks — mrjonfinger · 2026-09-21
- Andrew Chen: strong LLMs are far from running on phones, on-device AI faces bandwidth, heat and model-size hurdles — andrewchen · 2026-09-21