Skip to content
Back to Dwarkesh Podcast
Dwarkesh Podcast artwork
Indexed 0 mentions
Dwarkesh PodcastSep 26, 2025

Richard Sutton – Father of RL thinks LLMs are a dead end

Summary, transcript quotes, and episode notes for Richard Sutton – Father of RL thinks LLMs are a dead end on Dwarkesh Podcast.

Listen
Loading the embedded player…
Context before you listen

Richard Sutton – Father of RL thinks LLMs are a dead end is indexed here with episode context, audio, and transcript-derived mentions.

Episode summary
Richard Sutton is the father of reinforcement learning, winner of the 2024 Turing Award, and author of The Bitter Lesson. And he thinks LLMs are a dead end. After interviewing him, my steel man of Richard’s position is this: LLMs aren’t capable of learning on-the-job, so no matter how much we scale, we’ll need some new architecture to enable continual learning. And once we have it, we won’t need a special training phase — the agent will just learn on-the-fly, like all humans, and indeed, like all animals. This new paradigm will render our current approach with LLMs obsolete. In our interview, I did my best to represent the view that LLMs might function as the foundation on which experiential learning can happen… Some sparks flew. A big thanks to the Alberta Machine Intelligence Institute for inviting me up to Edmonton and for letting me use their studio and equipment. Enjoy! Watch on YouTube ; listen on Apple Podcasts or Spotify . Sponsors * Labelbox makes it possible to train AI agents in hyperrealistic RL environments. With an experienced team of applied researchers and a massive network of subject-matter experts, Labelbox ensures your training reflects important, real-world nuance. Turn your demo projects into working systems at labelbox.com/dwarkesh * Gemini Deep Research is designed for thorough exploration of hard topics. For this episode, it helped me trace reinforcement learning from early policy gradients up to current-day methods, combining clear explanations with curated examples. Try it out yourself at gemini.google.com * Hudson River Trading doesn’t silo their teams. Instead, HRT researchers openly trade ideas and share strategy code in a mono-repo. This means you’re able to learn at incredible speed and your contributions have impact across the entire firm. Find open roles at hudsonrivertrading.com/dwarkesh Timestamps (00:00:00) – Are LLMs a dead end? (00:13:04) – Do humans do imitation learning? (00:23:10) – The Era of Experience (00:33:39) – Current architectures generalize poorly out of distribution (00:41:29) – Surprises in the AI field (00:46:41) – Will The Bitter Lesson still apply post AGI? (00:53:48) – Succession to AIs Get full access to Dwarkesh Podcast at www.dwarkesh.com/subscribe
Book mentions0
Media mentions0
Quick answers

Quick FAQ

Answers to common summary, books, and takeaway questions for this episode.

What is Richard Sutton – Father of RL thinks LLMs are a dead end about?

Summary, transcript quotes, and episode notes for Richard Sutton – Father of RL thinks LLMs are a dead end on Dwarkesh Podcast.

What are the main takeaways from Richard Sutton – Father of RL thinks LLMs are a dead end?

These are the strongest takeaways surfaced by the transcript, summary copy, and linked mentions for Richard Sutton – Father of RL thinks LLMs are a dead end.

  • Richard Sutton is the father of reinforcement learning, winner of the 2024 Turing Award, and author of The Bitter Lesson. And he thinks LLMs are a dead end. After interv…

Which books are mentioned in Richard Sutton – Father of RL thinks LLMs are a dead end?

This episode does not have extracted book mentions yet, but the page still captures the core summary, audio, and transcript context.

Why are listeners searching for Richard Sutton – Father of RL thinks LLMs are a dead end?

Richard Sutton – Father of RL thinks LLMs are a dead end keeps attracting summary-style searches because this page combines episode context, transcript quotes, and direct jump links back into the audio.

Books Mentioned

The full list below is ranked by how useful each mention is to a listener: stronger recommendation language, clearer quote context, and better timestamp support rise first.

No book mentions yet

This episode does not have extracted book mentions yet.

Browse book directory
Weekly source-backed picks

Get the strongest books from new Dwarkesh Podcast episodes.

A short weekly email with transcript-backed book recommendations, source quotes, and exact moments from recently indexed episodes.

One useful email a week. Unsubscribe anytime.

Movies & Documentaries Mentioned

No movie or documentary mentions yet

This episode does not have extracted media mentions yet.