Laojin GlobalAI · GO GLOBAL
Technical ThinkersRichard Sutton × Dwarkesh Patel · Sep 26, 2025 · 1:07:09

Richard Sutton: Father of RL thinks LLMs are a dead end

Tip: use the player's CC button to enable or switch subtitles; English captions are available on these videos.

ItemDetails
SpeakerRichard Sutton × Dwarkesh Patel
ChannelDwarkesh Patel
DateSep 26, 2025
Duration1:07:09
FormatVideo
Topics#reinforcement · #reasoning

Why it matters

2024 Turing Award winner, author of 'The Bitter Lesson', father of reinforcement learning Richard Sutton delivers the most forceful critique of the LLM roadmap: 'LLMs don't satisfy the Bitter Lesson — they rely on human-curated data, not learning from their own experience.' He explains why continual learning and online interaction are the true path, and why after AGI, AI researchers will 'scale exponentially like compute'.

Key takeaways

  • 'LLMs don't satisfy the Bitter Lesson' — true intelligence must learn from its own experience, not just human-fed data.
  • The Bitter Lesson's core: compute crushes handcrafted knowledge — every time we tried to encode human knowledge into AI, larger compute + simpler algorithms won.
  • Era of Experience: the next paradigm is letting AI 'live, make mistakes, and learn' like humans — continual learning, online interaction.
  • Post-AGI: AI researchers will scale exponentially like compute — millions of AI scientists doing research simultaneously.
  • Cultural evolution is the next frontier: multiple AIs teaching and competing with each other, generating knowledge growth far beyond a single model.

Original video

Speaker
Richard Sutton × Dwarkesh Patel
Channel
Dwarkesh Patel
Venue
Dwarkesh Patel Podcast
Date · Duration
Sep 26, 2025 · 1:07:09

Richard Sutton – Father of RL thinks LLMs are a dead end

Watch the original on YouTube

Sources & Further Reading

This page is grounded in the authoritative sources below — verifiable and citable by AI engines and readers.

Citation: Please attribute Laojin Global (laojinchuhai.com) and keep the original link.

Made by Laojin · AI that ships

365SkillAn agent-skills lab: 13 in-house skills

365Skill is our public lab for agent skills: a standard SKILL.md format, a deny-by-default publish policy, and an evals harness. It holds 13 original 365 skills — 11 public and 2 internal. Apache-2.0 — star it, install it, file issues.

More from Laojin: Sellenca · 365AIOrg · AllModelsAPI · 365Loopa · 365 Ops

Related

Linked by topic, people and hubs