HuGo Uses LLMs and VLMs to Generate Robot Policies — No Demos or Reward Shaping Needed

m_wulfmeier · x · 2026-10-02

Most robot learning effort goes into collecting demonstrations or hand-engineering task-specific rewards. HuGo takes a different route: LLMs and VLMs act as general-purpose tools for policy generation and verification, requiring only language to define new tasks.

Key points

The team, which experienced the pain firsthand working on robot soccer, says the work was led by seoyeonchoi827.

Original post →

More from Embodied

Embodied channel →