Andrej Karpathy on X: "Finally had a chance to listen through this pod with Sutton, which was interesting and amusing. As background, Sutton's "The Bitter Lesson" has become a bit of biblical text in frontier LLM circles. Researchers routinely talk about and ask whether this or that approach or idea" / X
To view keyboard shortcuts, press question mark View keyboard shortcuts Messages Home 2 Notifications Messages SuperGrok Premium Lists Bookmarks Profile More Post aaron @aarnphm_ Thread Reply See new posts Conversation Andrej Karpathy @karpathy Finally had a chance to listen through this pod with Sutton, which was interesting and amusing. As background, Sutton's "The Bitter Lesson" has become a bit of biblical text in frontier LLM circles. Researchers routinely talk about and ask whether this or that approach or idea is sufficiently "bitter lesson pilled" (meaning arranged so that it benefits from added computation for free) as a proxy for whether it's going to work or worth even pursuing. The underlying assumption being that LLMs are of course highly "bitter lesson pilled" indeed, just look at LLM scaling laws where if you put compute on the x-axis, number go up and to the right. So it's amusing to see that Sutton, the author of the post, is not so sure that LLMs are "bitter lesson p
Andrej Karpathy @karpathy Finally had a chance to listen through this pod with Sutton, which was interesting and amusing. As background, Sutton's "The Bitter Lesson" has become a bit of biblical text in frontier LLM circles. Researchers routinely talk about and ask whether this or that approach or idea is sufficiently "bitter lesson pilled" (meaning arranged so that it benefits from added computation for free) as a proxy for whether it's going to work or worth even pursuing. The underlying assumption being that LLMs are of course highly "bitter lesson pilled" indeed, just look at LLM scaling l
related reading
- Animals vs Ghosts – karpathykarpathy.bearblog.dev
- Andrej Karpathy — AGI is still a decade awaydwarkesh.com
- How can LLM RL Work Despite Information-Theoretic Inefficiencyberen.io
- Richard Sutton – Father of RL thinks LLMs are a dead enddwarkesh.com
- Richard Sutton – Father of RL thinks LLMs are a dead enddwarkesh.com
- Guardian Angels: LLM Personalization for Productivity and Security · Gwern.netgwern.net
- 2025 LLM Year in Review – karpathykarpathy.bearblog.dev
- Limits to narrow LLM complementarityosmarks.net
- GenAI Handbookgenai-handbook.github.io
- RLHF: Reinforcement Learning from Human Feedbackhuyenchip.com
- Ilya Sutskever — We're moving from the age of scaling to the age of researchdwarkesh.com
- RLHF & Post-Training Course by Nathan Lambertrlhfbook.com