Reinforcement Learning
36 video summaries
Build a research pod on Reinforcement Learning.
36 videos curated. Save them to your own pod, ask any question across the body of expert opinion, and connect it to Claude or ChatGPT.
Videos About Reinforcement Learning

How Dopamine & Serotonin Shape Decisions, Motivation & Learning | Dr. Read Montague
Andrew Huberman

⚡️Factorio Learning Environment: the ultimate Game Agent Eval — Jack Hopkins
Latent Space

⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect
Latent Space

The #1 SWE-Bench Verified Agent
Latent Space

How Claude Plays Pokémon was made
Latent Space

Can We Contain Artificial Intelligence?: A Conversation with Mustafa Suleyman (Episode #332)
Sam Harris

Stanford Robotics Seminar ENGR319 | Spring 2026 | Interactive Autonomy
Stanford Online

5 Papers That Show Where AI Research Is Heading Right Now
Y Combinator

⚡️Every product of the future will be a living system — Ronak Malde, Trajectory.ai
Latent Space

Stanford MS&E435 Economics of the AI Supercycle | Spring 2026 | Enterprise Internal Knowledge
Stanford Online

The Key Thing Human Brains Have That AI Is Trying To Learn
Y Combinator

Stanford CS547 HCI Seminar | Spring 2026 | Promoting Agency in Human-AI Interaction
Stanford Online

Cooking with OpenAI’s Research Chief: AGI, o1, Evals, and Scaling Laws — Mark Chen
Latent Space

Stanford CS336 Language Modeling from Scratch | Spring 2026 | Lecture 16: Post-Training - RLVR
Stanford Online

Poolside’s Model Factory, Laguna S, Open Models, and the Race to AGI — Eiso Kant, Poolside AI
Latent Space

Building Dota Bots That Beat Pros - OpenAI's Greg Brockman, Szymon Sidor, and Sam Altman
Y Combinator

An AI Primer with Wojciech Zaremba
Y Combinator

Is AI Already Conscious? (FULL EPISODE)
Sam Harris

Stanford CS229 Machine Learning | Spring 2026 | Lecture 20: GMM (EM), PCA
Stanford Online

Stanford CS229 Machine Learning | Spring 2026 | Lecture 18: GMM (EM), PCA
Stanford Online

The Engineering Unlocks Behind DeepSeek | YC Decoded
Y Combinator

Stanford CS25: Transformers United V6 I From Next-Token Prediction to Next-Generation Intelligence
Stanford Online

Sergey Levine: Robotics and Machine Learning | Lex Fridman Podcast #108
Lex Fridman

Matt Botvinick: Neuroscience, Psychology, and AI at DeepMind | Lex Fridman Podcast #106
Lex Fridman
Page 1 of 2Next