Optimizing for efficiency with IBM’s Granite episode artwork

EPISODE · Mar 14, 2025 · 43 MIN

Optimizing for efficiency with IBM’s Granite

from Practical AI · host Daniel Whitenack and Chris Benson

We often judge AI models by leaderboard scores, but what if efficiency matters more? Kate Soule from IBM joins us to discuss how Granite AI is rethinking AI at the edge—breaking tasks into smaller, efficient components and co-designing models with hardware. She also shares why AI should prioritize efficiency frontiers over incremental benchmark gains and how seamless model routing can optimize performance. Featuring:Kate Soule – LinkedInChris Benson – Website, GitHub, LinkedIn, XDaniel Whitenack – Website, GitHub, XLinks:IBM GraniteIBM Granite on Hugging FaceIBM Expands Granite Model Family with New Multi-Modal and Reasoning AI Built for the Enterprise 

Episode metadata supplied by the publisher feed · Published Mar 14, 2025

Embed this episode

Ready to play

Optimizing for efficiency with IBM’s Granite

0:00 43:38

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of Practical AI?

This episode is 43 minutes long.

When was this Practical AI episode published?

This episode was published on March 14, 2025.

Is there a transcript available for this episode?

Yes, a full transcript is available for this episode. You can read the complete transcript on the episode page.

Can I download this Practical AI episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!