Skild releases S1, a robot foundation model that learns tasks from video demos

deepakpathak · x · 2026-08-26

Skild has released S1, a robot foundation model that takes task specifications via video demonstrations rather than language instructions. S1 handles unseen, long-horizon tasks up to 10 minutes with dozens of steps (e.g., potting plants, pour-over coffee) and can recover from mistakes or imperfect demos. It translates visual demonstrations into its own body context, addressing the limitations of language for delicate or complex tasks.

Related event: Skild AI Unveils S1, a Robot Foundation Model That Learns New Tasks from a Single Video(9 posts)→

Original post →

More from Embodied

Embodied channel →