Home Podcasts TalkRL: The Reinforcement Learning Podcast
TalkRL: The Reinforcement Learning Podcast

TalkRL: The Reinforcement Learning Podcast

Robin Ranjit Singh Chauhan 74 Episodes Aug 14, 2026

TalkRL is a podcast dedicated to reinforcement learning, featuring in-depth interviews with researchers and practitioners at the forefront of the field. Guests come from major institutions and labs such as MILA, OpenAI, MIT, DeepMind, Berkeley, Amii, Oxford, Google Research, Brown, Waymo, Caltech, and the Vector Institute. The show explores both fundamental ideas and practical applications in RL. It is hosted by Robin Ranjit Singh Chauhan.

Episodes

Thomas Frost on Clinical RL with Natural Timings
Thomas Frost on Clinical RL with Natural Timings Aug 14, 2026 5244 Dr Thomas Frost is an emergency physician based in London, UK. He is also in the final stages of completing a PhD at University College London, where he has been looking at offline reinforcement learning applied to healthcare settings.Featured ReferencesRobust Real-Time Mortality Prediction in the Intensive Care Unit using Temporal Difference Learning Thomas Frost, Kezhi Li, Steve Harris
Danijar Hafner on Dreamer v4
Danijar Hafner on Dreamer v4 Nov 9, 2025 6052 Danijar Hafner was a Research Scientist at Google DeepMind until recently.Featured References   Training Agents Inside of Scalable World Models [ blog ]  Danijar Hafner, Wilson Yan, Timothy LillicrapOne Step Diffusion via Shortcut ModelsKevin Frans, Danijar Hafner, Sergey Levine, Pieter AbbeelAction and Perception as Divergence Minimization [ blog ] Danijar Hafner, Pedro A. Ortega, Jimmy
David Abel on the Science of Agency @ RLDM 2025
David Abel on the Science of Agency @ RLDM 2025 Sep 8, 2025 3582 David Abel is a Senior Research Scientist at DeepMind on the Agency team, and an Honorary Fellow at the University of Edinburgh. His research blends computer science and philosophy, exploring foundational questions about reinforcement learning, definitions, and the nature of agency.  Featured References  Plasticity as the Mirror of Empowerment   David Abel, Michael Bowling, André Barreto,
Jake Beck, Alex Goldie, & Cornelius Braun on Sutton's OaK, Metalearning, LLMs, Squirrels @ RLC 2025
Jake Beck, Alex Goldie, & Cornelius Braun on Sutton's OaK, Metalearning, LLMs, Squirrels @ RLC 2025 Aug 19, 2025 740 Recorded at Reinforcement Learning Conference 2025 at University of Alberta, Edmonton Alberta Canada.Featured ReferencesLecture on the Oak Architecture, Rich SuttonAlberta Plan, Rich Sutton with Mike Bowling and Patrick Pilarski Additional ReferencesJacob Beck on Google Scholar Alex Goldie on Google ScholarCornelius Braun on Google ScholarReinforcement Learning Conference
Outstanding Paper Award Winners - 2/2 @ RLC 2025
Outstanding Paper Award Winners - 2/2 @ RLC 2025 Aug 17, 2025 858 We caught up with the RLC Outstanding Paper award winners for your listening pleasure. Recorded on location at Reinforcement Learning Conference 2025, at University of Alberta, in Edmonton Alberta Canada in August 2025.Featured References Empirical Reinforcement Learning ResearchMitigating Suboptimality of Deterministic Policy Gradients in Complex Q-functionsAyush Jain, Norio Kosaka, Xinh
Outstanding Paper Award Winners - 1/2 @ RLC 2025
Outstanding Paper Award Winners - 1/2 @ RLC 2025 Aug 15, 2025 406 We caught up with the RLC Outstanding Paper award winners for your listening pleasure.  Recorded on location at Reinforcement Learning Conference 2025, at University of Alberta, in Edmonton Alberta Canada in August 2025.Featured References  Scientific Understanding in Reinforcement Learning  How Should We Meta-Learn Reinforcement Learning Algorithms?  Alexander David Goldie, Zilin Wang, J
Thomas Akam on Model-based RL in the Brain
Thomas Akam on Model-based RL in the Brain Aug 3, 2025 3126 Prof Thomas Akam is a Neuroscientist at the Oxford University Department of Experimental Psychology.  He is a Wellcome Career Development Fellow and Associate Professor at the University of Oxford, and leads the Cognitive Circuits research group.Featured ReferencesBrain Architecture for Adaptive BehaviourThomas Akam, RLDM 2025 TutorialAdditional ReferencesThomas Akam on Google ScholarpyPh
Stefano Albrecht on Multi-Agent RL @ RLDM 2025
Stefano Albrecht on Multi-Agent RL @ RLDM 2025 Jul 22, 2025 1894 Stefano V. Albrecht was previously Associate Professor at the University of Edinburgh, and is currently serving as Director of AI at startup Deepflow. He is a Program Chair of RLDM 2025 and is co-author of the MIT Press textbook "Multi-Agent Reinforcement Learning: Foundations and Modern Approaches".Featured ReferencesMulti-Agent Reinforcement Learning: Foundations and Modern ApproachesSt
Satinder Singh: The Origin Story of RLDM @ RLDM 2025
Satinder Singh: The Origin Story of RLDM @ RLDM 2025 Jun 25, 2025 357 Professor Satinder Singh of Google DeepMind and U of Michigan is co-founder of RLDM.  Here he narrates the origin story of the Reinforcement Learning and Decision Making meeting (not conference).Recorded on location at Trinity College Dublin, Ireland during RLDM 2025.Featured ReferencesRLDM 2025: Multi-disciplinary Conference on Reinforcement Learning and Decision Making (RLDM)June 11-14,
NeurIPS 2024 - Posters and Hallways 3
NeurIPS 2024 - Posters and Hallways 3 Mar 9, 2025 601 Posters and Hallway episodes are short interviews and poster summaries.  Recorded at NeurIPS 2024 in Vancouver BC Canada.   Featuring  Claire Bizon Monroc from Inria: WFCRL: A Multi-Agent Reinforcement Learning Benchmark for Wind Farm Control  Andrew Wagenmaker from UC Berkeley: Overcoming the Sim-to-Real Gap: Leveraging Simulation to Learn to Explore for Real-World RL  Harley Wiltzer fro
NeurIPS 2024 - Posters and Hallways 2
NeurIPS 2024 - Posters and Hallways 2 Mar 4, 2025 528 Posters and Hallway episodes are short interviews and poster summaries.  Recorded at NeurIPS 2024 in Vancouver BC Canada.   Featuring  Jonathan Cook from University of Oxford: Artificial Generational Intelligence: Cultural Accumulation in Reinforcement Learning  Yifei Zhou from Berkeley AI Research: DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning
NeurIPS 2024 - Posters and Hallways 1
NeurIPS 2024 - Posters and Hallways 1 Mar 2, 2025 572 Posters and Hallway episodes are short interviews and poster summaries.  Recorded at NeurIPS 2024 in Vancouver BC Canada.   Featuring  Jiaheng Hu of University of Texas: Disentangled Unsupervised Skill Discovery for Efficient Hierarchical Reinforcement Learning  Skander Moalla of EPFL: No Representation, No Trust: Connecting Representation, Collapse, and Trust Issues in PPO  Adil Zouitine

Recommended