Why Fei-Fei Li Is Betting on Spatial Intelligence

a16z · 2026-09-04 · 44 min
https://www.youtube.com/watch?v=qn1QDDBnTA0Video summary
Fei-Fei Li and World Labs unveil Atlas, using new-view prediction to cut 3D capture from hundreds of images to just three.
World Labs co-founders Fei-Fei Li, Justin Johnson, and Ben Mildenhall explain Atlas, a world model that generates, reconstructs, and simulates scenes through “new view prediction.” Unlike video models that predict the next frame, Atlas uses spatial context and camera poses to produce grounded views, joining pixel generation and 3D reconstruction in one multimodal model. The team says a Matrix-style bullet-time shot can be made with three iPhone cameras instead of hundreds, while room capture may drop from 100–300 images to three. Atlas builds on lessons from Marble, whose Gaussian-splat output constrained the system. The founders see uses in creative production, architecture, and robotics, where real-to-sim data collection is costly. They say Atlas already supports some motion, but stronger dynamics, editing controls, and action planning remain future goals; training compute is the main limit to scaling.
Chapters
- 0:00What Atlas Is & Why It Matters: New-View Prediction and Bullet Time from Three Cameras
- 5:15Is This a Scaled-Up Video Model or a New Architecture? Atlas Unifies Generation and Reconstruction
- 8:15Spatial Intelligence & Why New View Prediction Matters: Atlas Moves Beyond Marble’s Gaussian Splats
- 13:34Spatial Intelligence & Why New View Prediction Matters: Atlas Targets 50–100× Fewer Capture Images
- 17:05Spatial Intelligence & Why New View Prediction Matters: Atlas Reconstructs Stanford Quad from Ground Views
- 21:27Did You Know It Was Going to Work? Scaling Conviction and the NeRF Table Breakthrough
- 24:42Use Cases: Creatives, Games & Robotics—Marvel’s Gaussian Splats and Virtual Design for Architecture
- 28:05Use Cases: Creatives, Games & Robotics—Atlas Speeds Synnex’s Real-to-Sim Robotics Data Pipeline
- 32:17Use Cases: Creatives, Games & Robotics—Learned Simulators Could Train Policies and Become Planners
- 35:21The Elephant in the Room: Video Models vs World Models—Atlas Adds Dynamics Beyond Static Marble
- 37:55Will We Get 4D Video You Can Walk Around In?—Editability, Scene Controls, and Spatial Interaction
- 42:13Why New View Prediction Is the Next Token Prediction—Atlas’s Generative Viewpoint Primitive
This is a Tier 1 public summary
Whether the chapter key points, section summaries and mind map are public is up to the person who shared it. Want the full analysis?Submit one yourself.
More from this channel
How Cursor Built One of AI’s Fastest-Growing Companiesa16zCursor bet on the human-model interface, grew rapidly without early sales hires, and later expanded through Graphite and founder acquisitions.
Why Top Founders Are Racing Into AI Infrastructurea16za16z’s Machine Age Fund targets AI infrastructure as GPU supply is booked to 2028 and memory demand needs three years of capacity.
Why AI Demand Is Outrunning Compute Supplya16zGavin Baker argues AI compute demand will outpace supply, with sub-one-year infrastructure paybacks and orbital data centers approaching.
What Today’s Best Models Still Can’t Do in Matha16zDaniel Litt praises AI’s Erdős unit-distance result but warns that proofs alone cannot replace mathematical understanding or human curiosity.
Inside Moderna’s Biggest mRNA Test Since COVIDa16zModerna and Merck’s personalized mRNA vaccine beat Keytruda alone in a Phase 3 melanoma trial, after 1,000 cancer-vaccine trials failed.
Related analyses
Why Physical AI Is the Next Big Opportunity | Deep Dives with a16za16z Deep DivesDiode Computers aims to automate circuit-board design in two years, while Unlimited Industries targets end-to-end construction automation within a decade.
Why Robotics Still Isn't Solved - But Could Be Soon | YC Paper ClubY CombinatorYC researchers say robotics remains unsolved, but memory, self-supervised reasoning, and simulation-trained dexterity are pushing it forward.
Chelsea Finn: This is the State of the Art in RoboticsY CombinatorChelsea Finn says Physical Intelligence’s PIO7 controls diverse robots out of the box, while reinforcement learning doubled throughput.
Going In Deep On Data | YC Paper ClubY CombinatorYC’s data talks show how expert-built benchmarks, diffusion models reaching 1,000 tokens per second, and multilingual scaling can improve AI.
Why Investors Are Rethinking Everything for the AI Eraa16za16z’s David George and Accolade’s Aram Verdiyan argue AI is intensifying venture power laws, with only 20 of 3,000 US firms delivering consistent 3x net returns.