Pure RL discovers superhuman robot strategies in sim, transfers zero-shot to real hardware

KyleMorgenstein · x · 2026-10-11

RL researcher robleerl amplified a demo showing robots behaving "like robots should": octizhang reports that pure RL, trained only in simulation, discovered strategies no human demonstrator would teach — the robot tips a rod upright by pushing it against a board leg, stands a gear by leveraging its shaft, and flips a nut in place instead of regrasping. All strategies suit the robot's own gripper rather than a human hand, and transferred zero-shot to the real robot. The takeaway: RL can unlock the true potential of a given embodiment beyond human priors.

Original post →

More from Embodied

Embodied channel →