EPISODE · Jul 22, 2020 · 26 MIN
Give that model a treat! : Reinforcement learning explained
from Tic-Tac-Toe the Hard Way · host People + AI Research
Switching gears, we focus on how Yannick’s been training his model using reinforcement learning. He explains the differences from David’s supervised learning approach. We find out how his system performs against a player that makes random tic-tac-toe moves.Resources: Deep Learning for JavaScript bookPlaying Atari with Deep Reinforcement LearningTwo Minute Papers episode on Atari DQNFor more information about the show, check out pair.withgoogle.com/thehardway/.You can reach out to the hosts on Twitter: @dweinberger and @tafsiri.
Embed this episode
What this episode covers
Switching gears, we focus on how Yannick’s been training his model using reinforcement learning. He explains the differences from David’s supervised learning approach. We find out how his system performs against a player that makes random tic-tac-toe moves.
NOW PLAYING
Give that model a treat! : Reinforcement learning explained
No transcript for this episode yet
Similar Episodes
No similar episodes found.
Similar Podcasts
No similar podcasts found.