

Richard Sutton – Father of RL thinks LLMs are a dead end
Richard Sutton is the father of reinforcement learning, winner of the 2024 Turing Award, and author of The Bitter Lesson. And he thinks LLMs are a dead end. After interviewing him, my steel man of Richard’s position is this: LLMs aren’t capable of learning on-the-job, so no matter how much we scale, we’ll need some…
We score an episode from what we can actually measure — its reach and engagement, what listeners say about it, and what it covers. We don't have those signals for this one yet, so it doesn't get a number.
Buzzmeter says it's buzzing across platforms — the Hive Vote says whether the people who actually listened liked it.
Rate this episode
Add your vote to the Hive — listeners rate every episode after they finish.
- No reviews yet — be the first.
More from Dwarkesh Podcast
See all episodes →Similar episodes from other shows
More like Dwarkesh Podcast →
#170 Richard Sutton on Pursuing AGI Through Reinforcement Learning

[AIEWF Preview] Multi-Turn RL for Multi-Hour Agents — with Will Brown, Prime Intellect

EP 41: The Reward Signal: The Missing Ingredient in Every AI System You’ve Built

Are World Models the Key to AGI?
One great episode in your inbox, daily — free.
One email a day, unsubscribe anytime. No spam, ever.


