System 3 thinking – Educating Silicon
I’ve been thinking for a while that there’s a piece missing from LLMs. There are hints that this hole might soon be filled, and it could drive the next leg up in AI capabilities. Many people have observed that LLMs, for all their abilities, seem to lack “spark”. The new reasoning models are remarkably good at a certain kind of knowledge-based problem solving, based on chaining together obscure facts, but they don’t seem to show the novel creative insights that characterize top human solutions. It’s somewhat reminiscent of the Deep Blue era in computer chess: the models approach problems in a grind-it-out kind of way. Humans sometimes do this too, but also have some other mode which the models seem to lack. Will this just fall out of further scaling? Or do we need some new ideas? While I am very bullish on scaling, I also think ideas are going to matter. In humans, it’s pretty well accepted that there are two distinct modes of cognition: System 1 and System 2. System 1 is basically asso
I’ve been thinking for a while that there’s a piece missing from LLMs. There are hints that this hole might soon be filled, and it could drive the next leg up in AI capabilities. Many people have observed that LLMs, for all their abilities, seem to lack “spark”. The new reasoning models are remarkably good at a certain kind of knowledge-based problem solving , based on chaining together obscure facts, but they don’t seem to show the novel creative insights that characterize top human solutions. It’s somewhat reminiscent of the Deep Blue era in computer chess: the models approach problems
Explore this link on the map →related reading
- LLM Daydreaming · Gwern.netgwern.net
- Andrej Karpathy — AGI is still a decade awaydwarkesh.com
- DeepSeek-R1arxiv.org
- the-illusion-of-thinking.pdfml-site.cdn-apple.com
- AI in 2025: gestalt — LessWronglesswrong.com
- GenAI Handbookgenai-handbook.github.io
- There Are No New Ideas in AI… Only New Datasetsblog.jxmo.io
- As Rocks May Think | Eric Jangevjang.com
- Language Models in Plato's Cave - by Sergey Levinesergeylevine.substack.com
- Generative AI's Act o1: The Reasoning Era Begins | Sequoia Capitalsequoiacap.com
- Optimizing LLM Test-Time Compute Involves Solving a Meta-RL Problem – Machine Learning Blog | ML@CMU | Carnegie Mellon Universityblog.ml.cmu.edu
- I am worried about near-term non-LLM AI developments — LessWronglesswrong.com