Agents & InferenceHacker News

World Labs unveils Atlas, a multimodal world model for 3D generation and simulation

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

World Labs' Atlas is a from-scratch multimodal autoregressive diffusion transformer that takes text, images, video, and 3D into a unified spatial context, generating 3D-consistent novel views and up to 1-minute 1440p videos from as few as one to six reference images with explicit camera-geometry control rather than text prompts. If you're building spatial/robotics simulation or 3D content pipelines, this shifts you from stochastic prompt-and-pray generation to deterministic camera-path staging, giving you reproducible view synthesis you can actually integrate into planning and rendering workflows.

LinkedIn

Two AI summaries of each story, blind-voted — see today's agents & inference digest →