Ant LingBot-Depth2.0 Tops 12 World Rankings, Open-Sources Embodied AI Vision Base
新智元 · wechat · 2026-07-07
Ant LingBot released the LingBot-Depth2.0 spatial perception model, securing 12 global first places across public and private datasets. It tackles core embodied AI challenges like depth estimation for glass, mirrors, and transparent objects, alongside fine object detection and long-range perception. Abandoning general vision models like DINOv3, the team built LingBot-Vision from scratch—the world's first "spatial-native" vision foundation model for embodied AI. With just 1.1 billion parameters, it outperforms 7B-level competitors. LingBot-Vision is open-sourced on Hugging Face, ModelScope, and GitHub, with its technical report on arxiv. Commercially, Ant LingBot partnered with Orbbec to launch the EGO-RGBD data collection device with integrated SDK, completing the loop from model capability to commercial deployment.
Related event: Ant Robbyant Open-Sources LingBot Vision Models, Topping Depth Benchmarks(18 posts)→
More from Embodied
- Nothing phone mockup turns a film joke into a modular design meme — ZeYanjie · 2026-07-22
- Lightwheel AI Launches SimReadyGen: Text-to-Physics-Accurate Robot Sim Assets — ZeYanjie · 2026-07-22
- Humanoid robot sorting packages in a warehouse sparks debate over job loss — MonaJalal_ · 2026-07-22
- NVIDIA pushes OpenUSD as the common layer for simulation and physical AI — MonaJalal_ · 2026-07-22
- A quadruped robot gets a custom glow-up with a new shell and screen — DynamicWebPaige · 2026-07-22
- A VR teleop demo for an SO-101 arm gets absurdly low latency by using one Python script — MoonL88537 · 2026-07-22