UpNext AI Deep Dive: World Models, Spatial Intelligence, and the Race to Teach AI Reality episode artwork

EPISODE · May 23, 2026 · 21 MIN

UpNext AI Deep Dive: World Models, Spatial Intelligence, and the Race to Teach AI Reality

from UpNext AI · host UpNext Labs

In this deep-dive episode of UpNext AI, we explore the growing debate around world models — AI systems designed to predict and reason about how the world changes over time. Large language models made AI useful as a software and knowledge interface, but researchers like Yann LeCun and Fei-Fei Li argue that acting in the physical world requires something more: spatial understanding, prediction, planning, and a model of consequences.We break down why world models are attracting major investment, how they differ from traditional robotics, why video models changed the conversation, and what recent research papers suggest about the path from passive observation to real-world action. We also look at the risks: unclear architectures, expensive data, reliability gaps, and the challenge of turning compelling research into durable businesses.Sources and further readingInterviewsFei-Fei Li interview: https://youtu.be/wDeXfFQcJxk?si=9oxB3NWXZiqeuj1KYann LeCun interview: https://youtu.be/_PioN-CpOP0?si=K7RRD7BtfKpQ9cCICompany and funding contextReuters — Fei-Fei Li’s World Labs raises $1 billion in funding: https://www.reuters.com/business/ai-pioneer-fei-fei-lis-world-labs-raises-1-billion-funding-2026-02-18/World Labs — funding announcement: https://www.worldlabs.ai/blog/funding-2026TechCrunch — Yann LeCun’s AMI Labs raises $1.03 billion to build world models: https://techcrunch.com/2026/03/09/yann-lecuns-ami-labs-raises-1-03-billion-to-build-world-models/Research papersV-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning: https://arxiv.org/abs/2506.09985Humanoid World Models: Open World Foundation Models for Humanoid Robotics: https://arxiv.org/abs/2506.01182GenCast: Probabilistic Weather Forecasting with Machine Learning: https://www.nature.com/articles/s41586-024-08252-9WorldSimBench / Towards Video Generation Models as World Simulators: https://openreview.net/forum?id=ejGAytoWoeVideo models and robotics contextOpenAI — Video generation models as world simulators: https://openai.com/index/video-generation-models-as-world-simulators/OpenAI — Sora: Creating video from text: https://openai.com/index/sora/Boston Dynamics — Large Behavior Models and Atlas Find New Footing: https://bostondynamics.com/blog/large-behavior-models-atlas-find-new-footing/Toyota Research Institute — AI-Powered Robot by Boston Dynamics and TRI takes key step toward general-purpose humanoids: https://www.tri.global/news/ai-powered-robot-boston-dynamics-and-toyota-research-institute-takes-key-step-towards-generalIEEE Spectrum — Boston Dynamics Atlas Learns From Large Behavior Models: https://spectrum.ieee.org/boston-dynamics-atlas-scott-kuindersmaAudio generation: ElevenLabs https://www.upnext.fm/eleven

Episode metadata supplied by the publisher feed · Published May 23, 2026

Embed this episode

NOW PLAYING

UpNext AI Deep Dive: World Models, Spatial Intelligence, and the Race to Teach AI Reality

0:00 21:17

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of UpNext AI?

This episode is 21 minutes long.

When was this UpNext AI episode published?

This episode was published on May 23, 2026.

Can I download this UpNext AI episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!