AlphaGo
Media
Used as an example of successful reinforcement learning in domains with clear win-loss conditions.
Mentioned in 3 videos
Save the 3 videos on AlphaGo to your own pod.
Sign up free to keep building your knowledge base on AlphaGo as more episodes are added.
Videos Mentioning AlphaGo

Stanford CS336 Language Modeling from Scratch | Spring 2026 | Lecture 16: Post-Training - RLVR
Stanford Online
Used as an example of successful reinforcement learning in domains with clear win-loss conditions.

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 18: RL Policy Optimization
Stanford Online
A project that demonstrated the success of reinforcement learning, particularly actor-critic methods, in playing the game of Go.

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 17: RL Value-Based Methods
Stanford Online
A famous example of reinforcement learning success in games, highlighting the potential for learning in high-dimensional state spaces.