Large models on CPUs (Practical AI #221) episode artwork

EPISODE · May 2, 2023 · 38 MIN

Large models on CPUs (Practical AI #221)

from Changelog Master Feed

Model sizes are crazy these days with billions and billions of parameters. As Mark Kurtz explains in this episode, this makes inference slow and expensive despite the fact that up to 90%+ of the parameters don't influence the outputs at all. Mark helps us understand all of the practicalities and progress that is being made in model optimization and CPU inference, including the increasing opportunities to run LLMs and other Generative AI models on commodity hardware.

Episode metadata supplied by the publisher feed · Published May 2, 2023

Embed this episode

NOW PLAYING

Large models on CPUs (Practical AI #221)

0:00 38:30

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Changelog Master Feed?

This episode is 38 minutes long.

When was this Changelog Master Feed episode published?

This episode was published on May 2, 2023.

Is there a transcript available for this episode?

Yes, a full transcript is available for this episode. You can read the complete transcript on the episode page.

Can I download this Changelog Master Feed episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!