Robot Long Context Scales to 8k Steps
ihorbeaver · x · 2026-07-18
The author discussed a piece of work on RoboTTT: long-horizon visuo-motor context is crucial for robot foundation models, but naively adding context significantly increases robot inference time, making it unscalable to very long time spans.
By integrating Test-Time-Training into the foundation model, this method saves context into fast weights that can be updated in constant time during inference. This scales the context size up to 8k timesteps without a noticeable increase in inference overhead. The author believes this could be a great fit for the MicroFactory stack.
More from Embodied
- Tesla expands Robotaxi rides to seven areas, including new Orlando and Tampa zones — elonmusk · 2026-07-22
- Hands-on robotics workshop on Saturday may be the last in-person session before August — StewartalsopIII · 2026-07-22
- NVIDIA pitches World Foundation Models as a way to scale physical AI data generation — MonaJalal_ · 2026-07-22
- RoboMME Podcast Preview: Benchmarking Memory for Robotic Policies — chris_j_paxton · 2026-07-21
- Gritt says an 8-person crew now installs 3,000 to 4,000 solar panels a day — HaktanSuren · 2026-07-21
- A helium-powered flying robot whale aims to be a quiet companion pet — chris_j_paxton · 2026-07-21