DeepSeek-R1: Redefining AI Reasoning with Pure Reinforcement Learning episode artwork

EPISODE · Sep 19, 2025 · 11 MIN

DeepSeek-R1: Redefining AI Reasoning with Pure Reinforcement Learning

from The Deep Dive Lab: Unraveling Materials Science · host Son Hoang

Explore how DeepSeek-R1, a groundbreaking Chinese LLM, leverages the Group Relative Policy Optimization (GRPO) framework to master advanced reasoning in math and coding. With low training costs and open weights, this Nature-published model is reshaping global AI research.

Episode metadata supplied by the publisher feed · Published Sep 19, 2025

Embed this episode

NOW PLAYING

DeepSeek-R1: Redefining AI Reasoning with Pure Reinforcement Learning

0:00 11:25

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of The Deep Dive Lab: Unraveling Materials Science?

This episode is 11 minutes long.

When was this The Deep Dive Lab: Unraveling Materials Science episode published?

This episode was published on September 19, 2025.

Can I download this The Deep Dive Lab: Unraveling Materials Science episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!