Open-Source Embodied Video Foundation Model
nikola_mr64990 · x · 2026-07-09
Robbyant has open-sourced LingBot-Video, an embodied intelligence video foundation model utilizing a MoE architecture with 30B total parameters, of which only 3B are activated during inference. The model builds on large-scale internet video pre-training, supplemented with 70,000 hours of embodied data. The post claims it outperforms Wan2.6, Seedance 1.5 Pro, and Cosmos3 Super on RBench, aiming to accelerate the deployment of embodied applications like robotics.
Related event: Ant Group Open-Sources LingBot-Video for Embodied AI(26 posts)→
More from Embodied
- Amazon and Google sold 600M+ smart speakers, so why no AGI-era successor? — julianlehr · 2026-09-11
- ECCV26 Oral: Flow Matching Enables Single-Stage Multi-View Point Cloud Registration — ducha_aiki · 2026-09-11
- Polish developers build iPhone app that detects nearby Meta smart glasses — Low-Honeydew6483 · 2026-09-11
- Ant's Afu health AI hits 150M users, unveils AI+hardware health alliance at Bund Summit — APPSO · 2026-09-11
- Johns Hopkins Launches Full-Stack Hands-on Robot Learning Class with SO-101 Arm Kits — _krishna_murthy · 2026-09-11
- SyncWorld: In-Context Robot World Model Simulates Unseen Views and Embodiments Zero-Shot — ChongZzZhang · 2026-09-11