LingBot-Video Training: 6D Physical Rewards and Real Robot Data
thetripathi58 · x · 2026-07-09
LingBot-Video changes the traditional scoring mechanism of video models that only pursue visual aesthetics. It adopts a single-step GRPO algorithm and introduces 6 precise reward signals: visual quality, image-text alignment, dynamics, motion coherence, human action consistency, and physical plausibility, ensuring the model understands real physical laws.
To prevent the model from "faking" physics, the team trained it using over 70,000 hours of real robot data, including real-hand operations, walking trajectories, and first-person perspectives from humanoid and quadruped robots.
Related event: Ant Group Open-Sources LingBot-Video for Embodied AI(26 posts)→
More from Embodied
- Tesla expands Robotaxi rides to seven areas, including new Orlando and Tampa zones — elonmusk · 2026-07-22
- Hands-on robotics workshop on Saturday may be the last in-person session before August — StewartalsopIII · 2026-07-22
- NVIDIA pitches World Foundation Models as a way to scale physical AI data generation — MonaJalal_ · 2026-07-22
- RoboMME Podcast Preview: Benchmarking Memory for Robotic Policies — chris_j_paxton · 2026-07-21
- Gritt says an 8-person crew now installs 3,000 to 4,000 solar panels a day — HaktanSuren · 2026-07-21
- A helium-powered flying robot whale aims to be a quiet companion pet — chris_j_paxton · 2026-07-21