RoboTTT Enables Long-Memory Robot Learning

DrJimFan · x · 2026-07-15

Introducing RoboTTT: applying test-time training to robotic policies. By performing gradient updates on a small kernel network during inference, it compresses historical experience into the model weights.

Key results include:

The authors also report a new context scaling curve: from 128 to 8K timesteps, closed-loop performance consistently improves without saturating; 8K context pre-training boosts performance by 62% compared to 1K.

Related event: RoboTTT brings test-time training to robot policies(6 posts)→

Original post →

More from Embodied

Embodied channel →