Training Data
Training Data
Sequoia Capital
Rich Sutton and Khurram Javed: Why AI Models Stop Learning, and How to Start It Again
53 minutes Posted Aug 18, 2026 at 9:00 am.
Introduction
An AI winter, a cancer diagnosis, and the move to Alberta
Writing "The Bitter Lesson," and what people get wrong
Are LLMs a positive or a negative example of it?
Synthetic data is "just a big mistake," and the Big World Hypothesis
AlphaGo, human priors, and why prior knowledge and learning should be friends
"Their weights never change": do LLM assistants actually learn?
Babies, squirrels, and why no animal learns by supervised learning
Rockets, imagination, and where paradigm shifts come from
The Alberta Plan and its 12 steps
Catastrophic forgetting and the cure
Oak's biggest ambition: a self-maintaining mind
Why the big labs are stuck in a local minimum
If everything goes right: LLMs, many minds, and hiring
0:00
53:43
Download MP3
Show notes
Rich Sutton, who helped pioneer reinforcement learning and wrote the seminal AI essay The Bitter Lesson, has now cofounded Oak Lab with his former student Khurram Javed. Their goal: to build agents that continuously learn from their own experience rather than from us. Rich doesn't think he holds a radical view: "I'm not weird. The field is weird." He says all learning is continual, and the field is the one that needed a new name for it. Rich and Khurram argue synthetic data is "a big mistake." Their "big world hypothesis" is that the world is massively more complex than any agent or simulator, so approximations have to be updated continuously rather than frozen at deployment. Rich calls LLMs an unanticipated scientific breakthrough, but says they represent roughly a quarter of intelligence. He says catastrophic forgetting is "totally curable" with the ideas behind their continual backprop algorithm. Khurram explains why the frontier labs can't follow: they sit in a local minimum where a new paradigm gets worse before it gets better. Their target, five to ten years out, is a trillion-parameter mind that keeps learning, stays coherent, and runs on 20 watts.
Hosted by Sonya Huang and Alfred Lin, Sequoia Capital