ImageNet
A large-scale visual database that was a pivotal inflection point for big data in AI, enabling breakthroughs in computer vision.
What podcasters actually say about ImageNet.
38 mentions, no marketing. Save them all to a pod and ask any question.
Common Themes
Videos Mentioning ImageNet

MIT 6.S094: Computer Vision
Lex Fridman
A large-scale image dataset with millions of images and thousands of categories, widely used for training and evaluating computer vision models.

MIT Sloan: Intro to Machine Learning (in 360/VR)
Lex Fridman
A dataset used for image recognition tasks, cited in discussions about how neural networks can be easily fooled and for comparing machine vs. human performance.

MIT 6.S094: Convolutional Neural Networks for End-to-End Learning of the Driving Task
Lex Fridman
One of the largest fully labeled image datasets, used for hierarchical category classification and for training state-of-the-art CNNs.

MIT 6.S094: Deep Reinforcement Learning for Motion Planning
Lex Fridman
A large dataset of labeled images used for training and evaluating machine learning models, particularly for image recognition tasks.

Andrej Karpathy: Tesla AI, Self-Driving, Optimus, Aliens, and AGI | Lex Fridman Podcast #333
Lex Fridman
A large visual database designed for use in visual object recognition software research, significant for enabling the deep learning revolution but now considered 'crushed' like MNIST for main research.

Stephen Wolfram: Fundamental Theory of Physics, Life, and the Universe | Lex Fridman Podcast #124
Lex Fridman
A large visual database designed for use in visual object recognition software research.

Robots Are Finally Starting to Work
Y Combinator
A benchmark dataset that significantly impacted the vision community, used as a comparison to the Open-X dataset in terms of scale and impact.

๐ฌ Training Transformers to solve 95% failure rate of Cancer Trials โ Ron Alfa & Daniel Bear, Noetik
Latent Space
A large, curated image dataset that was crucial for the advancement of deep learning in computer vision.

Recursion Is The Next Scaling Law In AI
Y Combinator
A large-scale image dataset used for training and benchmarking computer vision models.

Full AI Prompting Course with Andrew Ng
DeepLearningAI
A state-of-the-art video generation model by Google from 2022, mentioned to show the significant progress in AI video generation over a few years, contrasting its earlier artifacts with modern capabilities.

Stanford CS229 Machine Learning | Spring 2026 | Lecture 13: LLMs, Next-Word Prediction Loss
Stanford Online
A large-scale dataset used for training image classification models, often featuring around a thousand labels, which can be used for supervised pre-training to learn representations.

Stanford CS229 Machine Learning | Spring 2026 | Lecture 6: Dataset Split, ML Advice
Stanford Online
A popular dataset used for benchmarking computer vision models. The lecture discusses a rebuilt version of ImageNet to test for adaptive overfitting and found that model rankings remained consistent despite shifts in accuracy, suggesting overfitting might be less of an issue than previously thought.

Stanford CS229 Machine Learning | Spring 2026 | Lecture 1: Introduction
Stanford Online
A large dataset of 1 million image and label pairs collected by Fei-Fei Li's team, crucial for the advancement of image classification.

Using AI to Increase Your Intelligence & Enrich Humanity | Dr. Fei-Fei Li
Andrew Huberman
A large-scale dataset of 15 million images, collected and led by Fei-Fei Li's lab, instrumental in driving the development of machine learning algorithms for object recognition.