Robot One-Shot Imitation: From Specialized Architectures to Generalist Models

animesh_garg · x · 2026-08-30

A robotics expert compares technical paths from 2020 to today: achieving one-shot imitation previously required carefully designed specialized neural architectures (like Neural Task Graphs). Now, generalist foundation models (e.g., GEN-1.5, Skild S1) enable physical in-context learning simply by inputting video prompts directly into context. While scaling has improved convenience, the author notes a surprising fact: performance rates for one-shot imitation haven't drastically improved compared to the specialized methods achieved in 2020.

Related event: Robot One-Shot Imitation: Bigger Models, Little Real Progress(2 posts)→

Original post →

More from Embodied

Embodied channel →