LingBot-Video: 30B Params with 3B Active for Embodied Video AI

alifcoder · x · 2026-07-21

LingBot-Video focuses on improving the efficiency of large video models rather than simply scaling them up. Its flagship MoE model has 30B total parameters but only activates 3B during generation.

This approach offers an interesting direction for embodied AI video models by leveraging MoE architecture for better efficiency.

Original post →

More from Multimodal

Multimodal channel →