LM101-062: How to Transform a Supervised Learning Machine into a Value Function Reinforcement Learning Machine episode artwork

EPISODE · Mar 19, 2017 · 31 MIN

LM101-062: How to Transform a Supervised Learning Machine into a Value Function Reinforcement Learning Machine

from Learning Machines 101

This 62nd episode of Learning Machines 101 (www.learningmachines101.com)  discusses how to design reinforcement learning machines using your knowledge of how to build supervised learning machines! Specifically, we focus on Value Function Reinforcement Learning Machines which estimate the unobservable total penalty associated with an episode when only the beginning of the episode is observable. This estimated Value Function can then be used by the learning machine to select a particular action in a given situation to minimize the total future penalties that will be received. Applications include: building your own robot, building your own automatic aircraft lander, building your own automated stock market trading system, and building your own self-driving car!!

Episode metadata supplied by the publisher feed · Published Mar 19, 2017

Embed this episode

NOW PLAYING

LM101-062: How to Transform a Supervised Learning Machine into a Value Function Reinforcement Learning Machine

0:00 31:05

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of Learning Machines 101?

This episode is 31 minutes long.

When was this Learning Machines 101 episode published?

This episode was published on March 19, 2017.

Can I download this Learning Machines 101 episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!