Adapting Robot Models with Human Correction Data
micoolcho · x · 2026-07-16
This discussion highlights that while modern vision-language-action (VLA) and world-action models perform well in manipulation, they still struggle to reliably adapt when **transferred to new robots and tasks**. The proposed approach uses **DAgger-style online imitation learning**: - Deploy the robot first - Continuously collect human correction data - Update the policy dynamically, rather than relying solely on offline fine-tuning The author adds that if similar trajectories or human corrections are available, it would be valuable to scale up the accumulation of **DAgger-style data** across more tasks and environments.
Related event: FlowDAgger: Adapting Robot Models via Human Corrections(2 posts)→
More from Embodied
- BrainCo demos near-real-time bionics without implants and claims 85% lower prosthetic cost — TrueOrange9944 · 2026-07-21
- OpenAI’s $230 CodexMicro sold out, and users are already cloning it with Stream Decks — APPSO · 2026-07-21
- Xiaomi-Robotics-1 shows robot motion improves more from data than bigger models — The Decoder · 2026-07-21
- HarmoHOI generates multi-view hand-object videos and aligned 3D motion in one diffusion model — cn-scut · 2026-07-21
- Tesla is reportedly building a humanoid robot factory aimed at 10 million units a year — davidpattersonx · 2026-07-21
- Anthropic is reportedly in talks to buy Physical Intelligence — MarvinTBaumann · 2026-07-21