Pierluca D'Oro and Martin Klissarov episode artwork

EPISODE · Nov 13, 2023 · 57 MIN

Pierluca D'Oro and Martin Klissarov

from TalkRL: The Reinforcement Learning Podcast · host Robin Ranjit Singh Chauhan

Pierluca D'Oro and Martin Klissarov on Motif and RLAIF, Noisy Neighborhoods and Return Landscapes, and more!  Pierluca D'Oro is PhD student at Mila and visiting researcher at Meta.Martin Klissarov is a PhD student at Mila and McGill and research scientist intern at Meta.  Featured References  Motif: Intrinsic Motivation from Artificial Intelligence Feedback  Martin Klissarov*, Pierluca D'Oro*, Shagun Sodhani, Roberta Raileanu, Pierre-Luc Bacon, Pascal Vincent, Amy Zhang, Mikael Henaff  Policy Optimization in a Noisy Neighborhood: On Return Landscapes in Continuous Control  Nate Rahn*, Pierluca D'Oro*, Harley Wiltzer, Pierre-Luc Bacon, Marc G. Bellemare  To keep doing RL research, stop calling yourself an RL researcher Pierluca D'Oro 

Episode metadata supplied by the publisher feed · Published Nov 13, 2023

Embed this episode

Ready to play

Pierluca D'Oro and Martin Klissarov

0:00 57:24

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of TalkRL: The Reinforcement Learning Podcast?

This episode is 57 minutes long.

When was this TalkRL: The Reinforcement Learning Podcast episode published?

This episode was published on November 13, 2023.

Is there a transcript available for this episode?

Yes, a full transcript is available for this episode. You can read the complete transcript on the episode page.

Can I download this TalkRL: The Reinforcement Learning Podcast episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!