Fei-Fei Li's World Labs unveils Atlas, an omni world model turning 1-6 photos into navigable 3D worlds

APPSO · wechat · 2026-09-02

World Labs, founded by Fei-Fei Li, has released Atlas, an omni world model for spatial intelligence. Pretrained from scratch, it natively handles text, images, video, camera poses and 3D depth, unifying world generation, spatial reconstruction and spatiotemporal simulation in one model.

Key capabilities

Architecture: a multimodal autoregressive diffusion Transformer combining autoregressive spatial-state prediction, Rectified Flow denoising, and Transformer scaling, enabling KV cache, distributed inference and diffusion distillation reuse. In blind tests of camera control it beats MiniMax H3 75%, Gemini Omni Flash 81%, FLUX3 93% and Seedance2.5 94% of the time.

For robotics: Atlas targets embodied AI's sim-data bottleneck — traditional data collection is estimated to cost on the order of $100 trillion at scale. Two 24-frame phone videos can now yield a high-fidelity 3D physical space with simulated RGB, depth and force feedback. World Labs reports emergent abilities with scaling. Atlas is in Early Access for select partners and will underpin Marble and other products.

Related event: World Labs Launches Atlas, an Omni World Model for Spatial Intelligence(61 posts)→

Original post →

More from Embodied

Embodied channel →