← All stories
● Covered by 1 source · 1 reportMedium impact1 neutral

World Labs Introduces Atlas, a Multimodal World Model for Spatial Intelligence

🔄 Updated 2h ago
New to BrevFeed? We gather this story from every outlet covering it into one summary — ranked by real-world impact, not just the latest headline — so you never miss what matters. What is BrevFeed? →

Key points

  • Atlas is a multimodal autoregressive diffusion transformer.
  • It natively operates on text, images, video, and 3D inputs.
  • Capabilities include 1440p video generation and 3D scene reconstruction.
  • Atlas will power future World Labs products like Marble.

Introducing Atlas: A Next-Generation World Model

World Labs has unveiled Atlas, its latest world model designed for spatial intelligence. Atlas is an omni model, pretrained from scratch to process and generate content across text, images, video, and 3D formats. It functions as a multimodal autoregressive diffusion transformer, integrating all inputs into a unified spatial context to generate subsequent content while maintaining 3D consistency.

Core Capabilities and Features

Atlas offers a range of functionalities. It can perform camera-controlled generation, producing up to one minute of 1440p video from one or more images with precise camera control. For spatial reconstruction, Atlas reconstructs real-world scenes from multiple input images, generating novel views and explicit 3D outputs. The model also handles space-time simulation, reframing videos and enabling Real-to-Sim workflows for robotics. Additionally, Atlas generates images and 360 panoramas from text prompts, supporting complex instructions and various visual styles.

Technical Approach and Scalability

Atlas combines diverse inputs into a shared spatial context, using this context to predict and generate subsequent elements. The model maintains 3D consistency with observed data and extrapolates beyond it. World Labs states that Atlas is built for scalability, with performance improving as training compute increases, a trend expected to continue with further scaling efforts.

Applications and Future Impact

Atlas is intended to power future versions of World Labs' products, including Marble. Its ability to generate new views from reference images with specified camera positions and angles, while extrapolating scene content and geometry, makes it applicable across various scene types, visual styles, and camera motions. The model's native use of precise camera geometry as an input type allows for detailed shot framing and motion control.

✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →

The daily brief

One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.

One email a day. Unsubscribe in one click, any time.

Today's brief

Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.

~24 min · 20 stories · Sep 01

▶ Play today's brief Listen on Spotify

New every morning, and the back catalogue is archived by date.

Reporting from

World Labs has launched Atlas, a new world model pretrained on text, images, video, and 3D data. Atlas is designed to generate, reconstruct, and simulate worlds, offering capabilities like camera-controlled video generation, spatial reconstruction, and space-time simulation. This model aims to advance spatial intelligence for applications in creative content, high-fidelity simulations, and robotics.