SDS 551: Deep Reinforcement Learning — with Wah Loon Keng

Topics covered
Popular Clips
Episode Highlights
RL Strategies
The evolution of reinforcement learning (RL) strategies is marked by significant milestones, such as the development of AlphaGo Zero and AlphaZero. highlights how AlphaGo Zero learned without human data, a leap from its predecessor, AlphaGo, which relied on human gameplay. AlphaZero further advanced this by mastering multiple games like Go, chess, and shogi without human input 1.
AlphaZero is no human knowledge, but it masters go chess and another game called shogi.
---
The discussion also touches on the stagnation in new RL algorithms, with the last surge of innovations occurring in 2017. However, notes that DeepMind's work on open-ended play and emergent behavior offers promising directions for overcoming current limitations 2.
Implications
The long-term implications of reinforcement learning advances are profound, impacting both technology and society. explores how deep RL is currently applied in industries and speculates on its future potential. These advancements could revolutionize various sectors, enhancing automation and decision-making processes 3.
Very cool to hear about how deep reinforcement learning is being used in industry today.
---
As RL continues to evolve, it promises exciting developments in AI, potentially transforming how we interact with technology and solve complex problems 4.
Related Episodes

SDS 510: Deep Reinforcement Learning — with Jon Krohn
Answers 383 questions

SDS 503: Deep Reinforcement Learning for Robotics — with Pieter Abbeel
Answers 383 questions
SDS 464: A.I. vs Machine Learning vs Deep Learning — with Jon Krohn
Answers 383 questions

SDS 469: Learning Deep Learning Together — with Konrad Körding
Answers 383 questions

SDS 439: Deep Learning for Machine Vision — with Deblina Bhattacharjee
Answers 383 questions
SDS 558: @JonKrohnLearns's Answers to Questions on Machine Learning
Answers 383 questions
SDS 554: @JonKrohnLearns's Deep Learning Courses
Answers 383 questions

773: Deep Reinforcement Learning for Maximizing Profits — with Prof. Barrett Thomas
Answers 383 questions
SDS 506: Supervised vs Unsupervised Learning — with Jon Krohn
Answers 383 questions

797: Deep Learning Classics and Trends — with Dr. Rosanne Liu
Answers 383 questions

791: Reinforcement Learning from Human Feedback (RLHF) — with Dr. Nathan Lambert
Answers 383 questions

SDS 605: Upskilling in Data Science and Machine Learning — with Kian Katanforoosh
Answers 383 questions
SDS 438: Artificial General Intelligence — with Jon Krohn
Answers 383 questions
SDS 446: Getting Started in Machine Learning — with Jon Krohn
Answers 383 questions














