Storage Becomes The AI Bottleneck episode artwork

EPISODE · Mar 9, 2026 · 17 MIN

Storage Becomes The AI Bottleneck

from What's Up with Tech? · host Evan Kirstel

Interested in being a guest? Email us at [email protected] feels fast until memory and storage slow everything down. We sit with Michael Wu to unpack a blunt truth: inference is where value happens, and storage now sits on the critical path. Instead of treating SSDs as cold capacity, Phison’s adaptive middleware turns them into a live cache that expands usable memory and keeps models, embeddings, and long context windows close to compute. The payoff is practical and immediate—lean AIPCs and mini workstations run bigger workloads with steadier latency, and teams can scale inference without waiting for DRAM supply to catch up.We trace the story from CES announcements to real-world deployment. Michael breaks down how OEM integrations and consumer upgrade kits bring adaptive caching to both new and existing machines, why developer and education communities are the first winners, and how this bottom-up momentum seeds better software and on-device AI experiences. For enterprise leaders, we map the route from local experiments to global rollouts: consistent performance across distributed teams, lower cloud egress, and a storage layer tuned for retrieval-augmented generation, fine-tuning, and high-concurrency serving.Zooming out, we explore where we are in the AI cycle—early, hungry, and building—and how edge devices and “physical AI” will broaden demand for fast, cache-aware storage. Michael also shares Phison’s fabless strategy, the new Pascari enterprise lineup, and the push toward Gen 6 performance that aligns with next-gen model serving. If you care about real-world AI velocity, this conversation shows how to turn a bottleneck into an advantage by rethinking the memory hierarchy from the ground up.If this helped you think differently about scaling AI, follow the show, share it with a teammate, and leave a quick review so more builders can find it. What’s the first AI workflow you’d speed up with adaptive caching?Support the showMore at https://linktr.ee/EvanKirstel

Episode metadata supplied by the publisher feed · Published Mar 9, 2026

Embed this episode

Interested in being a guest? Email us at [email protected] AI feels fast until memory and storage slow everything down. We sit with Michael Wu to unpack a blunt truth: inference is where value happens, and storage now sits on the critical path. Instead of treating SSDs as cold capacity, Phison’s adaptive middleware turns them into a live cache that expands usable memory and keeps models, embeddings, and long context windows close to compute. The payoff is practical and immediate—lean AIPC...

Distinct summary based on available episode metadata or transcript content.

NOW PLAYING

Storage Becomes The AI Bottleneck

0:00 17:17

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of What's Up with Tech??

This episode is 17 minutes long.

When was this What's Up with Tech? episode published?

This episode was published on March 9, 2026.

Is there a transcript available for this episode?

Yes, a full transcript is available for this episode. You can read the complete transcript on the episode page.

Can I download this What's Up with Tech? episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!