Tic-Tac-Toe the Hard Way podcast artwork

PODCAST · technology

Tic-Tac-Toe the Hard Way

A writer and a software engineer from Google's People + AI Research team explore the human choices that shape machine learning systems by building competing tic-tac-toe agents.

Publisher-supplied feed metadata · PodParley refreshed Apr 9, 2026 · Source feed

  1. 10

    Lessons learned

    What have we learned about machine learning and the human decisions that shape it? And is machine learning perhaps changing our minds about how the world outside of machine learning — also known as the world — works?For more information about the show, check out pair.withgoogle.com/thehardway/.You can reach out to the hosts on Twitter: @dweinberger and @tafsiri. 

  2. 9

    Head to Head: The Even Bigger ML Smackdown!

    Yannick and David’s systems play against each other in 500 games. Who’s going to win? And what can we learn about how the ML may be working by thinking about the results?See the agents play each other in Tic-Tac-Two!For more information about the show, check out pair.withgoogle.com/thehardway/.You can reach out to the hosts on Twitter: @dweinberger and @tafsiri. 

  3. 8

    Enter tic-tac-two

    David’s variant of tic-tac-toe that we’re calling tic-tac-two is only slightly different but turns out to be far more complex. This requires rethinking what the ML system will need in order to learn how to play, and  how to represent that data.For more information about the show, check out pair.withgoogle.com/thehardway/.You can reach out to the hosts on Twitter: @dweinberger and @tafsiri. 

  4. 7

    Head to Head: the Big ML Smackdown!

    David and Yannick’s tic-tac-toe ML agents face-off against each other in tic-tac-toe!See the agents play each other!For more information about the show, check out pair.withgoogle.com/thehardway/.You can reach out to the hosts on Twitter: @dweinberger and @tafsiri. 

  5. 6

    Give that model a treat! : Reinforcement learning explained

    Switching gears, we focus on how Yannick’s been training his model using reinforcement learning.  He explains the differences from David’s supervised learning approach. We find out how his system performs against a player that makes random tic-tac-toe moves.Resources: Deep Learning for JavaScript bookPlaying Atari with Deep Reinforcement LearningTwo Minute Papers episode on Atari DQNFor more information about the show, check out pair.withgoogle.com/thehardway/.You can reach out to the hosts on Twitter: @dweinberger and @tafsiri. 

  6. 5

    Beating random: What it means to have trained a model

    David did it! He trained a machine learning model to play tic-tac-toe! (Well, with lots of help from Yannick.) How did the whole training experience go? How do you tell how training went? How did his model do against a player that makes random tic-tac-toe moves? For more information about the show, check out pair.withgoogle.com/thehardway/.You can reach out to the hosts on Twitter: @dweinberger and @tafsiri. 

  7. 4

    From tic-tac-toe moves to ML model

    Once we have the data we need—thousands of sample games--how do we turn it into something the ML can train itself on? That means understanding how training works, and what a model is. Resources:See a definition of one-hot encodingFor more information about the show, check out pair.withgoogle.com/thehardway.You can reach out to the hosts on Twitter: @dweinberger and @tafsiri. 

  8. 3

    What does a tic-tac-toe board look like to machine learning?

    How should David represent the data needed to train his machine learning system? What does a tic-tac-toe board “look” like to ML? Should he train it on games or on individual boards? How does this decision affect how and how well the machine will learn to play? Plus, an intro to reinforcement learning, the approach Yannick will be taking.For more information about the show, check out pair.withgoogle.com/thehardway.You can reach out to the hosts on Twitter: @dweinberger and @tafsiri. 

  9. 2

    Howdy, and the myth of “pouring in data”

    Welcome to the podcast! We’re Yannick and David, a software engineer and a non-technical writer. Over the next 9 episodes we’re going to use two different approaches to build machine learning systems that play two versions of tic-tac-toe. Building a machine learning app requires humans making a lot of decisions. We start by agreeing that David will use a “supervised learning” approach while Yannick will go with “reinforcement learning.”For more information about the show, check out pair.withgoogle.com/thehardway. You can reach out to the hosts on Twitter: @dweinberger and @tafsiri. 

  10. 1

    Introducing Tic-Tac-Toe the Hard Way

    Introducing the podcast where a writer and a software engineer explore the human choices that shape machine learning systems by building competing tic-tac-toe agents. Brought to you by Google's People + AI Research team.More at: pair.withgoogle.com/thehardway

Type above to search every episode's transcript for a word or phrase. Matches are scoped to this podcast.

Searching…

We're indexing this podcast's transcripts for the first time — this can take a minute or two. We'll show results as soon as they're ready.

No matches for "" in this podcast's transcripts.

Showing of matches

No topics indexed yet for this podcast.

Loading reviews...

ABOUT THIS SHOW

A writer and a software engineer from Google's People + AI Research team explore the human choices that shape machine learning systems by building competing tic-tac-toe agents.

HOSTED BY

People + AI Research

Produced by Lucas Dixon

CATEGORIES

Frequently Asked Questions

How many episodes does Tic-Tac-Toe the Hard Way have?

Tic-Tac-Toe the Hard Way currently has 10 episodes available on PodParley. New episodes are automatically indexed when they're published to the podcast feed.

What is Tic-Tac-Toe the Hard Way about?

A writer and a software engineer from Google's People + AI Research team explore the human choices that shape machine learning systems by building competing tic-tac-toe agents.

How often does Tic-Tac-Toe the Hard Way release new episodes?

Tic-Tac-Toe the Hard Way has 10 episodes. Check the episode list to see recent publication dates and frequency.

Where can I listen to Tic-Tac-Toe the Hard Way?

You can listen to Tic-Tac-Toe the Hard Way on PodParley by clicking any episode. We provide an embedded audio player for direct listening, and you can also subscribe via your preferred podcast app using the RSS feed.

Who hosts Tic-Tac-Toe the Hard Way?

Tic-Tac-Toe the Hard Way is created and hosted by People + AI Research.
URL copied to clipboard!