TANGO: whole-body VLA lets humanoid robots zero-shot navigate cluttered spaces on Unitree G1

chris_j_paxton · x · 2026-09-09

A CoRL 2026 paper introduces TANGO, a whole-body vision-language-action framework for humanoid navigation: from a language instruction and egocentric RGB, it directly predicts 29-DoF joint actions coordinating arms, torso, and gait to traverse cluttered 3D indoor spaces.

Related event: TANGO: Whole-Body VLA Enables Humanoid Robots to Navigate Cluttered Spaces(2 posts)→

Original post →

More from Embodied

Embodied channel →