RoboTTT: 8K Context Robot Policies

scott_e_reed · x · 2026-07-15

This research introduces RoboTTT, extending robot policies to an 8K timesteps visuo-motor context. It pushes the usable context length beyond current SOTA levels while maintaining constant inference latency.

Demonstrated capabilities include:

Comments emphasize that such TTT methods are highly suitable for robot learning and are likely agnostic to whether the backbone is a VLA or a video model.

Related event: RoboTTT brings test-time training to robot policies(6 posts)→

Original post →

More from Embodied

Embodied channel →