Atlas: A World Model for Spatial Intelligence

World Labs has introduced Atlas, a new 'world model' designed to simulate and reconstruct 3D environments from text, images, and video. The model uses spatial intelligence to generate consistent 3D scenes, allowing for precise camera control and creative scene generation.
Why it matters
Atlas represents a significant advancement in generative AI, moving beyond simple image generation toward complex, physically consistent 3D world simulation for robotics and creative media.
World models generate, reconstruct, and simulate any possible world. They understand how worlds appear, behave, and evolve so that we can render imagined worlds for creative users, simulate the real world in high fidelity, and help robots plan actions. At World Labs, we build these general purpose world models in pursuit of spatial intelligence.
Today we are introducing Atlas, our next-generation world model. Atlas is an omni model that we pretrained from scratch to natively operate on text, images, video, and 3D. It is a multimodal autoregressive diffusion transformer: all inputs are combined into a shared spatial context. Atlas uses that context to generate what comes next, staying consistent in 3D with everything it has seen and imagining what lies beyond it. Atlas is built to scale: its performance improves with increased training compute, and we expect this trend to hold as we continue scaling.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in