From Cloud Dependency to Local Intelligence: The Future of Accessible AI episode artwork

EPISODE · Oct 29, 2025 · 24 MIN

From Cloud Dependency to Local Intelligence: The Future of Accessible AI

from Intel on AI · host Intel Corporation

As AI models grow more powerful, the question of where they run is becoming just as important as what they do. In this episode, Brandon Weng, Co-Founder and CEO of Fluid Inference, unpacks what it takes to move AI from massive data centers to everyday devices—and why that shift matters. Brandon shares the story behind Fluid Inference, a company focused on making it easier for developers to deploy large AI models like transformers on consumer hardware. From pivoting away from his previous project, Slipbox, to the technical and philosophical choices that shaped Fluid's direction, he walks us through the thinking behind local-first AI. We explore the tradeoffs between cloud-based and on-device inference—touching on privacy, cost, control, and performance—and the hardware breakthroughs that are making edge AI more viable, including integrated NPUs in devices like Intel's AI PCs.   #EdgeAI #OnDeviceInference #AIOptimization #PrivacyFirst #OpenSourceAI #LocalAI

Episode metadata supplied by the publisher feed · Published Oct 29, 2025

Embed this episode

NOW PLAYING

From Cloud Dependency to Local Intelligence: The Future of Accessible AI

0:00 24:55

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Intel on AI?

This episode is 24 minutes long.

When was this Intel on AI episode published?

This episode was published on October 29, 2025.

Can I download this Intel on AI episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!