World Labs' Atlas rebuilds a full room from just 3 photos, a 50-100x capture reduction
drfeifei · x · 2026-09-05
World Labs co-founders Fei-Fei Li, Justin Johnson, and Ben Mildenhall sat down with a16z's Martin Casado to discuss Atlas, their world model for spatial intelligence.
- Core idea: LLMs are built on next-token prediction and video models on next-frame prediction; Atlas is built on new-view prediction, and is claimed to be the first model unifying pixel generation and pixel reconstruction—two tracks computer vision has kept separate for half a century.
- Practical result: Dense reconstruction of a room previously required 100-300 photos; Atlas works from just three, a 50-100x reduction in capture cost.
- Implication: That flip changes what imagery is worth reconstructing—old photos sitting in your camera roll can now be turned into 3D scenes. Mildenhall explains the key is finally reconciling data-driven priors with brute-force dense reconstruction.
Related event: Fei-Fei Li's World Labs unveils Atlas world model(3 posts)→
More from Models
- Astra 3D model floods X, but OpenAI missed the viral moment by delaying launch — bindureddy · 2026-09-05
- Qwen3.8 Max jumps 22% on new RSI-Exam benchmark for recursive self-improvement — cihangxie · 2026-09-05
- Stratechery: Anthropic walks back data retention policy, Nvidia earnings, Meta settles — Stratechery · 2026-09-05
- Claude Suddenly Replied in Russian to a User Who Never Spoke It — roshbakeer · 2026-09-05
- OpenAI's Astra uses 'recurrent depth' reasoning, obscuring its thinking process — JacquesThibs · 2026-09-05
- TAOCP open problems released as a dataset to benchmark frontier models — sytelus · 2026-09-05