The Future of Reasoning Models and AI Infrastructure episode artwork

EPISODE · Sep 30, 2025 · 32 MIN

The Future of Reasoning Models and AI Infrastructure

from Shared Everything · host Alon Horev, Kevin Deierling, Nicole Hemsoth Prickett

In this episode of the Shared Everything, reasoning models take center stage. No longer just text predictors, they now loop, branch, and drag in outside data, which blows open context windows and GPU limits. Alon Horev, CTO of VAST Data, unpacks how this shift strains infrastructure, while Kevin Deierling, SVP of Networking at NVIDIA, explains how NVIDIA Dynamo moves KV caches and workloads across GPUs, networks, and storage to keep agentic workflows moving. Data platforms become an extension of memory, enabling longer chains of thought, real-time agents, and secure, observable data paths. The result is a vivid picture of the AI datacenter as the nervous system for reasoning at scale.

Episode metadata supplied by the publisher feed · Published Sep 30, 2025

Embed this episode

NOW PLAYING

The Future of Reasoning Models and AI Infrastructure

0:00 32:29

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Shared Everything?

This episode is 32 minutes long.

When was this Shared Everything episode published?

This episode was published on September 30, 2025.

Can I download this Shared Everything episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!