World Labs releases Atlas, an omni world model native to text, images, video, and 3D

theworldlabs · x · 2026-09-24

World Labs has introduced Atlas, its next-generation omni world model for spatial intelligence, pretrained from scratch to natively operate on text, images, video, and 3D. Architecturally it is a multimodal autoregressive diffusion transformer: all inputs combine into a shared spatial context, and generation stays 3D-consistent with everything seen.

Key capabilities:

Atlas scales with training compute, a trend the team expects to hold. It will power future versions of Marble; beta signups are open.

Related event: Atlas Teases Chisel in Beta: Block Out a World and Let AI Bring It to Life(2 posts)→

Original post →

More from Multimodal

Multimodal channel →