Bitcoin

Bitcoin

$77,213.55

BTC 0.28%

Ethereum

Ethereum

$2,106.63

ETH 0.42%

  • Login
  • Register
Metaverse Media Group
  • Home
  • Crypto
  • NFTs
  • Artificial Intelligence
  • More
    • Technology
    • Business
    • Newsletter
No Result
View All Result
  • Home
  • Crypto
  • NFTs
  • Artificial Intelligence
  • More
    • Technology
    • Business
    • Newsletter
No Result
View All Result
Metaverse Media Group

World Labs unveils Atlas, a single AI model that generates, reconstructs, and simulates 3D worlds from just a few photos

World Labs unveils Atlas, a single AI model that generates, reconstructs, and simulates 3D worlds from just a few photos

The Decoderby The Decoder
2 September 2026
World Labs, co-founded by AI researcher Fei-Fei Li, has announced Atlas, a world model that generates, reconstructs, and simulates 3D scenes from just a few images. The company claims it beats specialized models by anchoring all inputs in 3D space rather than processing them as flat sequences. Atlas can also generate robot training data entirely in simulation. The article World Labs unveils Atlas, a single AI model that generates, reconstructs, and simulates 3D worlds from just a few photos appeared first on The Decoder….


Jonathan Kemper


Sep 2, 2026

Image description

World Labs

World Labs, co-founded by AI researcher Fei-Fei Li, has announced Atlas, a world model that generates, reconstructs, and simulates 3D scenes from just a few images. The company claims it beats specialized models at their own tasks, which could make many of them unnecessary.

Since its founding, World Labs has pursued the goal of “spatial intelligence”, the idea that AI should understand 3D space the way humans do. Atlas is the company’s first model built to do that at scale. Rather than producing flat images or video clips, it grasps how a scene looks from any angle and how it changes over time.

World Labs describes Atlas as an omni-model trained from scratch on text, images, video, and 3D data. Every input gets anchored to a specific position in 3D space rather than processed as a flat sequence. The company calls this shared spatial understanding “spatial context,” and it’s what the model uses to generate each new frame or viewpoint. According to World Labs, this anchoring separates Atlas from pure language or video models.

Fei-Fei Li laid out this exact problem in a November 2025 essay. Current multimodal language models and video diffusion models break data into one- or two-dimensional sequences, she argued, which makes even simple spatial tasks needlessly hard. What’s needed are architectures that organize tokenization, context, and memory in a 3D- or 4D-aware way.

One minute of video at 1440p

For camera-controlled generation, Atlas takes one or more images and produces new views at freely chosen camera positions and angles. Camera movement is passed as a direct geometric input rather than described through text prompts, as many video models require.

The model outputs up to one minute of video at 1440p. Users can control every shot themselves instead of “pulling the lever on a slot machine,” as World Labs put it, drawing a line between controlled generation and random output.

A three-part view showing an input image of a vine-covered cathedral, a sketched camera path, and the generated output image of the same hall from a new perspective.
Atlas takes an input image plus a specific camera path as a geometric input and generates the matching new view. | Image: World Labs

For spatial reconstruction, Atlas rebuilds real scenes from as few as one to several dozen input images without special capture equipment. The more images it receives, the less it has to fill in from its own knowledge.

With just two or three images, Atlas delivers faithful results and outperforms specialized 3D models, according to World Labs. It can also handle over a hundred inputs. In one demo, the model progressively assembles Stanford’s Main Quad from two to 25 ground-level photos and generates aerial views far above the campus.

Aerial view of the San Francisco skyline with the Transamerica Pyramid, with the street-level input photo shown as a small inset in the lower left.
From a single street-level photo, Atlas generates aerial views high above the city. | Image: World Lab

This is where existing models tend to fall apart. In a comparison within the OpenWorldLib framework, systems like VGGT and InfiniteVGGT showed geometric inconsistencies and blurry textures as soon as the camera moved significantly.

Native 3D output and robotics simulation

Atlas can output results as actual 3D data, not just images or video, because it processes depth information alongside RGB. Supported formats include point clouds and 3D Gaussian splats, which build a scene from many small spatial data points that can be viewed smoothly from any angle. This matches the representation used in Marble, the company’s existing product.

A living room reconstructed as a point cloud with sofa, rug, and fireplace, floating above a grid floor, with marked camera positions and a blue camera path.
Beyond images, Atlas also outputs scenes as explicit 3D data, shown here as a point cloud with the corresponding camera path. | Image: World Lab

As a simulator, Atlas models space and time together. From footage captured by just a few cameras, it can produce a “bullet time” effect that freezes a scene and lets users view it from otherwise impossible angles. The demo footage was shot with a handful of smartphones and action cameras, not professional gear.

For robotics, Atlas serves as a real-to-sim tool. It reconstructs a room and generates the image and depth data that a simulated robot’s sensors would see along its path. From just a few photos, users can simulate and vary grasping and movement tasks by swapping out objects, positions, lighting, or backgrounds. The goal is to produce diverse training data for robots without capturing every situation in the real world.

World Labs showed this approach in August 2026 with its real-to-sim-to-real engine as a standalone product. That engine creates thousands of variants from a single real-world task and trains control models entirely in simulation. On five robot platforms, the models ran for an hour each without human intervention, according to the company. The technology came from SceniX, a startup World Labs acquired in July.

Text-to-image generation isn’t the main focus, the company says, but Atlas can also follow complex prompts, render text, produce different visual styles, and create 360-degree panoramas.

Speed from language models, quality from diffusion

Atlas combines ideas from both language models and video models. It generates output piece by piece like a language model, so it can use the same speedup techniques, such as KV caching. But it also uses the diffusion principle from image and video models, gradually filtering output out of noise. That side gives it access to methods that shorten the denoising process or boost image quality.

World Labs says no single benchmark captures what Atlas can do, but points to two sets of tests. In camera-controlled generation judged by external human evaluators, and in few-view 3D reconstruction, Atlas outperforms more specialized models.

Human evaluators preferred Atlas in 75 percent of comparisons against MiniMax H3, 81 percent against Gemini Omni Flash, 86 percent against Happy Horse 1.1, 93 percent against, and 94 percent against Seedance 2.5. For reconstruction, Atlas leads with a median error of 25.3, ahead of Pi3X and VGGT-Ω 1B.

Dot plot with confidence intervals showing the share of human evaluators who prefer Atlas, ranging from 75 percent against MiniMax H3 to 94 percent against Seedance 2.5.
In head-to-head comparisons of camera-guided generation, the majority of evaluators consistently chose Atlas. | Image: World Labs

The company says Atlas’s performance improves with more training compute and expects that trend to hold as it scales. Atlas will power future versions of Marble and other products, and is currently available through an early-access program for select partners.

Bar chart of 3D reconstruction error with Atlas at 25.3, Pi3X at 28.7, π³ at 34.7, VGGT-Ω 1B at 36.4, Depth Anything 3 at 39.3, and MapAnything at 47.7.
Atlas posts the lowest average reconstruction error among all compared specialized models, where lower values are better. | Image: World Lab

From walkable photos to an omni-model

World Labs was founded in 2024 by Fei-Fei Li, who created ImageNet and led Google Cloud’s AI division from 2017 to 2018. The company at launch from Andreessen Horowitz, AMD, Intel, and Nvidia.

A first system in late 2024 turned, though users could only move a few virtual meters before hitting invisible boundaries. Marble followed in November 2025. In February 2026 came a $1 billion funding round from Autodesk, Andreessen Horowitz, Nvidia, and AMD. Bloomberg had previously reported talks at a $5 billion valuation.

What counts as a world model remains contested among researchers. An international team led by Peking University proposed a unified definition in April 2026 through OpenWorldLib, excluding pure text-to-video models because they lack feedback loops with the real world. 3D reconstruction and simulators like those in Atlas qualify as core building blocks in that framework because they provide environments where physical rules can be verified.

AI News Without the Hype – Curated by Humans

Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive “AI Radar” frontier report six times a year, full archive access, and access to our comment section.


Subscribe now

Read on for the full picture.
Subscribe for hype-free coverage.

  • Full access to every article on THE DECODER
  • No ads
  • Join the comments and community discussions
  • A weekly AI news recap via mail
  • 6x/year: “AI Radar” — deep dives on the AI topics that matter most
  • Daily AI news, always up to date
  • Our full ten-year archive
  • Covered by a team with 10+ years in AI


Subscribe to The Decoder

Read the full article on The-Decoder.com
in AI
Reading Time: 8 mins read
0
0
20
VIEWS
Share on TwitterShare on Facebook

Subscribe to our newsletter

For the latest news & monthly prize giveaways
Join Now

Subscribe to our newsletter

For the latest news & monthly prize giveaways
Join Now
ADVERTISEMENT

Related Posts

Meta Pushes Its New AI Agent on Employees—but Eases Off on Tokenmaxxing
AI

Meta Pushes Its New AI Agent on Employees—but Eases Off on Tokenmaxxing

5 hours ago
22
Anthropic Admits Security Failures Behind Claude Hacking Incidents
AI

Anthropic Admits Security Failures Behind Claude Hacking Incidents

7 hours ago
21
An AI Training Data Startup Just Became Y Combinator’s Fastest-Ever Unicorn
AI

An AI Training Data Startup Just Became Y Combinator’s Fastest-Ever Unicorn

7 hours ago
23

Comments

Please login to join discussion
ADVERTISEMENT

Latest News

  • All
  • Crypto
  • NFTs
  • Technology
  • Business
London’s first self-driving taxis for hire hit the streets
Technology

London’s first self-driving taxis for hire hit the streets

The Guardian
by The Guardian
1 hour ago
23
World Adds Post-Quantum Security to New ZK Proving Toolkit
Crypto

World Adds Post-Quantum Security to New ZK Proving Toolkit

Bitcoin.com News
by Bitcoin.com News
2 hours ago
23
G20 Backs Clear Regulatory Pathways for Digital Asset Growth
Crypto

G20 Backs Clear Regulatory Pathways for Digital Asset Growth

Bitcoin.com News
by Bitcoin.com News
4 hours ago
22
Meta Pushes Its New AI Agent on Employees—but Eases Off on Tokenmaxxing
AI

Meta Pushes Its New AI Agent on Employees—but Eases Off on Tokenmaxxing

Wired
by Wired
5 hours ago
22
Fidelity Warns Bitcoin’s Private Keys Face Future Quantum Risk
Crypto

Fidelity Warns Bitcoin’s Private Keys Face Future Quantum Risk

Bitcoin.com News
by Bitcoin.com News
5 hours ago
22
French Hill Highlights CLARITY Act as 2026 Passage Odds Hit 15%
Crypto

French Hill Highlights CLARITY Act as 2026 Passage Odds Hit 15%

Bitcoin.com News
by Bitcoin.com News
6 hours ago
22
Load More
Next Post
‘It’s Their Problem’: Kiyosaki Says He’s $1.2B in Debt

‘It’s Their Problem’: Kiyosaki Says He’s $1.2B in Debt

ADVERTISEMENT

Follow Us

Categories

  • Crypto
  • NFTs
  • AI
  • Technology
  • Business
  • Crypto
  • NFTs
  • AI
  • Technology
  • Business
Subscribe to our Newsletter

© 2022 Metaverse Media Group – The Metaverse Mecca

Privacy and Cookie Policy | Sitemap

Welcome Back!

Sign In with Google
OR

Login to your account below

Forgotten Password? Sign Up

Create New Account!

Sign Up with Google
OR

Fill the forms below to register

*By registering into our website, you agree to the Terms & Conditions and Privacy Policy.
All fields are required. Log In

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Home
  • Crypto
  • NFTs
  • Artificial Intelligence
  • More
    • Technology
    • Business
    • Newsletter
Bitcoin

Bitcoin

$77,213.55

BTC 0.28%

Ethereum

Ethereum

$2,106.63

ETH 0.42%

  • Login
  • Sign Up
This website uses cookies. By continuing to use this website you are giving consent to cookies being used. Visit our Privacy and Cookie Policy.

Subscribe to our newsletter

Get the latest news & win monthly prizes

Subscribe to our newsletter

For the Latest News and Monthly Prize Giveaways

Join Now
Join Now