Open-Sourced LingBot-VA 2.0 Robot Foundation Model
philfung · x · 2026-07-11
- Released LingBot-VA 2.0: A native video-action foundation model designed for general robot control.
- The author emphasizes it was pre-trained from scratch, unlike many works adapted from VLMs or pre-trained video models, and is fully open-source.
- Key designs include:
- Native video-action pre-training: Learns world knowledge to enhance generalization;
- Semantic Visual-Action Tokenizer: Improves action prediction accuracy and prompt-following capabilities;
- Foresight Reasoning: Enables the robot to "think one step ahead" during execution for smoother, more responsive control.
Related event: LingBot-VA/VLA 2.0 Released: Native Embodied Foundation Model(24 posts)→
More from Embodied
- The Humanoid AI raises $152M Series A at a $1.35B valuation — RazRazcle · 2026-07-22
- SceniX joins World Labs to close the real-to-sim gap for robot learning — davidyin44 · 2026-07-22
- Humanoid robot sorting packages in a warehouse sparks debate over job loss — MonaJalal_ · 2026-07-22
- NVIDIA pushes OpenUSD as the common layer for simulation and physical AI — MonaJalal_ · 2026-07-22
- A quadruped robot gets a custom glow-up with a new shell and screen — DynamicWebPaige · 2026-07-22
- NVIDIA says physical AI starts in simulation with OpenUSD and synthetic data — MonaJalal_ · 2026-07-22