// HACKER NEWS — CYBERSECURITY
Atlas: A World Model for Spatial Intelligence
World models generate, reconstruct, and simulate any possible world.
They understand how worlds appear, behave, and evolve
so that we can render imagined worlds for creative users,
simulate the real world in high fidelity, and help robots plan actions.
At World Labs, we build these general purpose world models in pursuit of spatial intelligence.
Today we are introducing Atlas, our next-generation world model.
Atlas is an omni model that we pretrained from scratch to natively operate on
text, images, video, and 3D.
It is a multimodal autoregressive diffusion transformer:
all inputs are combined into a shared spatial context.
Atlas uses that context to generate what comes next,
staying consistent in 3D with everything it has seen and imagining what lies beyond it.
Atlas is built to scale: its performance improves with increased training compute,
and we expect this trend to hold as we continue scaling.
Atlas can perform a broad range of tasks
spanning world generation, reconstruction, and simulation:
Atlas will power future versions of Marble
and other products from World Labs.
Atlas takes one or more reference images and generates
new views at any camera position and angle you specify.
Generated views match the content and geometry of the input images,
smoothly extrapolating beyond them to imagine parts of the scene
not visible in the inputs.
Atlas handles a broad range of scene types, visual styles, and camera motions.
Videos are generated from
one to six input images
with manually-designed camera paths
Atlas uses precise camera geometry as a native input type,
going beyond coarse text-based instructions for camera control.
This lets you frame every shot and control every motion.
In the examples here,
Atlas generates a complete scene from a single input image.
It uses the content of the input image along with its broad world knowledge
to imagine what the scene should look like from new angles.
For example, it generates the back side of the robot,
and it guesses that there should be a grassy lawn next to the pool.
From a single image, Atlas generates views from any angle. Drag to change
the view.