AIToday
Image GenerationAI Business & IndustrySiliconANGLE AIPublished: Sep 2, 2026, 13:00 JST2 min read

World Labs unveils Atlas, a 3D world model from one image

World Labs unveils Atlas, a 3D world model from one image

3 Key Points

  1. What happened

    World Labs, the AI startup co-founded by Fei-Fei Li, released Atlas, a multimodal world model that creates detailed 3D simulated environments from a single 2D image. It generates up to a minute of 1440p video with precise camera control and can output 3D assets like point clouds and Gaussian splats.

  2. Why it matters

    Atlas aims to bridge simulated environments and physical reasoning, fulfilling Li's theory of spatial intelligence—AI must reason about 3D objects, people, and interactions to understand the physical world. The startup has raised $1.2 billion from backers including Nvidia, AMD, and Autodesk, and sees robotics training as its real target.

  3. What to watch

    In blind tests, Atlas was overwhelmingly preferred over competitors like Gemini Omni Flash and FLUX for camera-path adherence. It is available in early access now for select enterprises, but no general release date has been announced.

Ask the AI about this article →

Summaries like this, in your inbox every morning.

Context & Analysis

World Labs' debut of Atlas marks a concrete step toward Fei-Fei Li's vision of spatial intelligence. Li, who led Stanford's AI Lab and co-founded the Stanford Institute for Human-Centered AI, launched World Labs in February 2024, arguing that AGI is impossible without understanding the physical world. Atlas, built on a multimodal autoregressive diffusion transformer architecture, goes beyond early video generators by treating camera trajectories and geometry as native inputs, not afterthoughts.

The startup's $1.2 billion in funding from Nvidia, AMD, and Autodesk shows strong investor confidence in this approach. However, Atlas faces a crowded field, with competitors like Odyssey, AMI Labs (founded by Yann LeCun), and Niantic Spatial each focusing on different aspects of world modeling. Atlas's bet is that a single base model can unify simulation, reconstruction, and camera control, offering a more integrated solution.

The real proof will come when wider availability lets users test whether early impressive results hold up in complex, real-world scenarios. Until then, the initial benchmark wins—especially the overwhelming preference over Gemini Omni Flash and FLUX for camera-path adherence—offer promising signs, but the market remains highly competitive.

FAQ
What makes Atlas different from traditional video generators?
Unlike text-prompt-based generators like Sora, Atlas ingests camera trajectories and geometry as native inputs, enabling precise perspective control. It also reconstructs scenes in 3D.
Who can use Atlas right now?
Atlas is available in early access for select enterprises. A general release date has not been announced.
SiliconANGLE AIRead Original Article

Also reported by ITmedia AI+

Get the latest Image Generation news every morning

For example, today's edition would include:

  • AI food ads spark backlash; restaurants say they're temporaryFortune AI · 1d ago
  • Why AI food images look so unappetizingThe Verge AI · 3d ago
  • AI-generated restaurant menus look off, and science explains whyTechCrunch AI · 3d ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleTCL CSOT bets on InP laser chips as supply tightens