Open-Source Embodied Video Foundation Model
nikola_mr64990 · x · 2026-07-09
Robbyant has open-sourced LingBot-Video, an embodied intelligence video foundation model utilizing a MoE architecture with 30B total parameters, of which only 3B are activated during inference. The model builds on large-scale internet video pre-training, supplemented with 70,000 hours of embodied data. The post claims it outperforms Wan2.6, Seedance 1.5 Pro, and Cosmos3 Super on RBench, aiming to accelerate the deployment of embodied applications like robotics.
Related event: Ant Group Open-Sources LingBot-Video for Embodied AI(26 posts)→
More from Embodied
- Tesla expands Robotaxi rides to seven areas, including new Orlando and Tampa zones — elonmusk · 2026-07-22
- Hands-on robotics workshop on Saturday may be the last in-person session before August — StewartalsopIII · 2026-07-22
- NVIDIA pitches World Foundation Models as a way to scale physical AI data generation — MonaJalal_ · 2026-07-22
- RoboMME Podcast Preview: Benchmarking Memory for Robotic Policies — chris_j_paxton · 2026-07-21
- Gritt says an 8-person crew now installs 3,000 to 4,000 solar panels a day — HaktanSuren · 2026-07-21
- A helium-powered flying robot whale aims to be a quiet companion pet — chris_j_paxton · 2026-07-21