LingBot-VA 2.0 Released

CyberRobooo · x · 2026-07-10

@robbyantbrain has released LingBot-VA 2.0, claiming it to be the first truly embodied-native foundation model. They emphasized that it was built from scratch for the physical world rather than being fine-tuned from a video generation model.

Key results and design specs include:

Architecturally, it features a semantic vision-action tokenizer, a causal DiT with sparse MoE, and a mechanism called Foresight Reasoning that predicts the next world state during robot motion. The author also noted that multiple components released this week collectively form a "closed-loop" embodied full-stack:

The overarching claim is that a single "brain" can serve multiple robots.

Related event: LingBot-VA/VLA 2.0 Released: Native Embodied Foundation Model(24 posts)→

Original post →

More from Embodied

Embodied channel →