Robot world model trained on 15 hours of video generalizes to unseen bodies

jon_barron · x · 2026-07-24

Researchers trained a video world model using only 15 hours of video from a single-arm robot.

The post highlights a compact but surprisingly general robot-world-model setup, with an emphasis on embodiment transfer and action generation.

Original post →

More from Embodied

Embodied channel →