Reinforcement Learning
45 video summaries
Build a research pod on Reinforcement Learning.
45 videos curated. Save them to your own pod, ask any question across the body of expert opinion, and connect it to Claude or ChatGPT.
Videos About Reinforcement Learning

How Dopamine & Serotonin Shape Decisions, Motivation & Learning | Dr. Read Montague
Andrew Huberman

⚡️Factorio Learning Environment: the ultimate Game Agent Eval — Jack Hopkins
Latent Space

⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
Latent Space

The #1 SWE-Bench Verified Agent
Latent Space

How Claude Plays Pokémon was made
Latent Space

Can We Contain Artificial Intelligence?: A Conversation with Mustafa Suleyman (Episode #332)
Sam Harris

Stanford Robotics Seminar ENGR319 | Spring 2026 | Interactive Autonomy
Stanford Online

5 Papers That Show Where AI Research Is Heading Right Now
Y Combinator

⚡️Every product of the future will be a living system — Ronak Malde, Trajectory.ai
Latent Space

Stanford MS&E435 Economics of the AI Supercycle | Spring 2026 | Enterprise Internal Knowledge
Stanford Online

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 1: Course Overview
Stanford Online

Stanford CS329A Self-Improving AI Agents | Part 6 | Train Time Scaling/Scaling RL
Stanford Online

The Key Thing Human Brains Have That AI Is Trying To Learn
Y Combinator

Stanford CS547 HCI Seminar | Spring 2026 | Promoting Agency in Human-AI Interaction
Stanford Online

Stanford CS229 Machine Learning | Spring 2026 | Lecture 1: Introduction
Stanford Online

Cooking with OpenAI’s Research Chief: AGI, o1, Evals, and Scaling Laws — Mark Chen
Latent Space

Stanford CS336 Language Modeling from Scratch | Spring 2026 | Lecture 16: Post-Training - RLVR
Stanford Online

Poolside’s Model Factory, Laguna S, Open Models, and the Race to AGI — Eiso Kant, Poolside AI
Latent Space

Chelsea Finn: This is the State of the Art in Robotics
Y Combinator

Building Dota Bots That Beat Pros - OpenAI's Greg Brockman, Szymon Sidor, and Sam Altman
Y Combinator

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 19: Model-Based RL
Stanford Online

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 15: Imitation Learning
Stanford Online

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 16: Fundamentals of RL
Stanford Online

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 18: RL Policy Optimization
Stanford Online
Page 1 of 2Next