Reinforcement Learning

45 video summaries

Build a research pod on Reinforcement Learning.

45 videos curated. Save them to your own pod, ask any question across the body of expert opinion, and connect it to Claude or ChatGPT.

Get Started Free

Videos About Reinforcement Learning

How Dopamine & Serotonin Shape Decisions, Motivation & Learning | Dr. Read Montague

How Dopamine & Serotonin Shape Decisions, Motivation & Learning | Dr. Read Montague

Andrew Huberman

⚡️Factorio Learning Environment: the ultimate Game Agent Eval — Jack Hopkins

⚡️Factorio Learning Environment: the ultimate Game Agent Eval — Jack Hopkins

Latent Space

⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect

⚡️Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect

Latent Space

The #1 SWE-Bench Verified Agent

The #1 SWE-Bench Verified Agent

Latent Space

How Claude Plays Pokémon was made

How Claude Plays Pokémon was made

Latent Space

Can We Contain Artificial Intelligence?: A Conversation with Mustafa Suleyman (Episode #332)

Can We Contain Artificial Intelligence?: A Conversation with Mustafa Suleyman (Episode #332)

Sam Harris

Stanford Robotics Seminar ENGR319 | Spring 2026 | Interactive Autonomy

Stanford Robotics Seminar ENGR319 | Spring 2026 | Interactive Autonomy

Stanford Online

5 Papers That Show Where AI Research Is Heading Right Now

5 Papers That Show Where AI Research Is Heading Right Now

Y Combinator

⚡️Every product of the future will be a living system  — Ronak Malde, Trajectory.ai

⚡️Every product of the future will be a living system — Ronak Malde, Trajectory.ai

Latent Space

Stanford MS&E435 Economics of the AI Supercycle | Spring 2026 | Enterprise Internal Knowledge

Stanford MS&E435 Economics of the AI Supercycle | Spring 2026 | Enterprise Internal Knowledge

Stanford Online

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 1: Course Overview

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 1: Course Overview

Stanford Online

Stanford CS329A Self-Improving AI Agents | Part 6 | Train Time Scaling/Scaling RL

Stanford CS329A Self-Improving AI Agents | Part 6 | Train Time Scaling/Scaling RL

Stanford Online

The Key Thing Human Brains Have That AI Is Trying To Learn

The Key Thing Human Brains Have That AI Is Trying To Learn

Y Combinator

Stanford CS547 HCI Seminar | Spring 2026 | Promoting Agency in Human-AI Interaction

Stanford CS547 HCI Seminar | Spring 2026 | Promoting Agency in Human-AI Interaction

Stanford Online

Stanford CS229 Machine Learning | Spring 2026 | Lecture 1: Introduction

Stanford CS229 Machine Learning | Spring 2026 | Lecture 1: Introduction

Stanford Online

Cooking with OpenAI’s Research Chief: AGI, o1, Evals, and Scaling Laws — Mark Chen

Cooking with OpenAI’s Research Chief: AGI, o1, Evals, and Scaling Laws — Mark Chen

Latent Space

Stanford CS336 Language Modeling from Scratch | Spring 2026 | Lecture 16: Post-Training - RLVR

Stanford CS336 Language Modeling from Scratch | Spring 2026 | Lecture 16: Post-Training - RLVR

Stanford Online

Poolside’s Model Factory, Laguna S, Open Models, and the Race to AGI — Eiso Kant, Poolside AI

Poolside’s Model Factory, Laguna S, Open Models, and the Race to AGI — Eiso Kant, Poolside AI

Latent Space

Chelsea Finn: This is the State of the Art in Robotics

Chelsea Finn: This is the State of the Art in Robotics

Y Combinator

Building Dota Bots That Beat Pros - OpenAI's Greg Brockman, Szymon Sidor, and Sam Altman

Building Dota Bots That Beat Pros - OpenAI's Greg Brockman, Szymon Sidor, and Sam Altman

Y Combinator

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 19: Model-Based RL

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 19: Model-Based RL

Stanford Online

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 15: Imitation Learning

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 15: Imitation Learning

Stanford Online

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 16: Fundamentals of RL

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 16: Fundamentals of RL

Stanford Online

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 18: RL Policy Optimization

Stanford AA203 Optimal and Learning-Based Control | Spring 2026 | Lecture 18: RL Policy Optimization

Stanford Online

Page 1 of 2Next