Robotics training debate: sim distillation vs. human behavior cloning

KyleMorgenstein · x · 2026-08-26

Kyle Morgenstein argues that distillation from single-task policies trained in simulation differs fundamentally from large-scale behavior cloning (BC) with human data, as single-task policies cannot scale the way BC + human data does. Chris Offner questions if multi-expert distillation represents a single policy performing multiple skills, noting that "athletic intelligence" remains capped at mimicking human motions or myopic MPC-like lookahead, failing to scale to other morphologies.

Related event: Debate: Sim Distillation vs Scalable Human Data in Robot Learning(2 posts)→

Original post →

More from Embodied

Embodied channel →