DeepSeek-R1: Reasoning-Driven AI with Reinforcement Learning episode artwork

EPISODE · Mar 3, 2025 · 19 MIN

DeepSeek-R1: Reasoning-Driven AI with Reinforcement Learning

from Ctrl Alt Society · host Halifax Studios

Today's deep dive, we talk about DeepSeek, a framework designed to enhance reasoning in large language models (LLMs) using reinforcement learning (RL), bypassing the need for extensive supervised fine-tuning. It introduces models like DeepSeek-R1-Zero and DeepSeek-R1, demonstrating emergent reasoning behaviors such as self-verification and reflection. 

Episode metadata supplied by the publisher feed · Published Mar 3, 2025

Embed this episode

NOW PLAYING

DeepSeek-R1: Reasoning-Driven AI with Reinforcement Learning

0:00 19:14

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of Ctrl Alt Society?

This episode is 19 minutes long.

When was this Ctrl Alt Society episode published?

This episode was published on March 3, 2025.

Can I download this Ctrl Alt Society episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!