Give that model a treat! : Reinforcement learning explained episode artwork

EPISODE · Jul 22, 2020 · 26 MIN

Give that model a treat! : Reinforcement learning explained

from Tic-Tac-Toe the Hard Way · host People + AI Research

Switching gears, we focus on how Yannick’s been training his model using reinforcement learning.  He explains the differences from David’s supervised learning approach. We find out how his system performs against a player that makes random tic-tac-toe moves.Resources: Deep Learning for JavaScript bookPlaying Atari with Deep Reinforcement LearningTwo Minute Papers episode on Atari DQNFor more information about the show, check out pair.withgoogle.com/thehardway/.You can reach out to the hosts on Twitter: @dweinberger and @tafsiri. 

Episode metadata supplied by the publisher feed · Published Jul 22, 2020

Embed this episode

Switching gears, we focus on how Yannick’s been training his model using reinforcement learning. He explains the differences from David’s supervised learning approach. We find out how his system performs against a player that makes random tic-tac-toe moves.

Distinct summary based on available episode metadata or transcript content.

NOW PLAYING

Give that model a treat! : Reinforcement learning explained

0:00 26:04

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Tic-Tac-Toe the Hard Way?

This episode is 26 minutes long.

When was this Tic-Tac-Toe the Hard Way episode published?

This episode was published on July 22, 2020.

Can I download this Tic-Tac-Toe the Hard Way episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!