✳flâneur — a map of the web's best reading
Why reasoning models will generalize - by Nathan Lambert
interconnects.ai · 1,952 words · saved by 1 readers
People underestimate the long-term potential of “reasoning.”
Why reasoning models will generalize People underestimate the long-term potential of “reasoning.” Nathan Lambert Jan 28, 2025 219 7 24 Share Article voiceover 0:00 -11:37 Audio playback is not supported on your browser. Please upgrade. This post is early to accommodate some last minute travel on my end! The new models trained to express extended chain of thought are going to generalize outside of their breakthrough domains of code and math. The “reasoning” process of language models that we use today is chain of thought reasoning. We ask the model to work step by step because it helps it manag
Explore this link on the map →related reading
- DeepSeek-R1arxiv.org
- the-illusion-of-thinking.pdfml-site.cdn-apple.com
- o1 and Reasoning | AndoLogsblog.ando.ai
- DeepSeek R1's recipe to replicate o1 and the future of reasoning LMsinterconnects.ai
- The State of Reinforcement Learning for LLM Reasoningmagazine.sebastianraschka.com
- Explore | alphaXivalphaxiv.org
- [2201.11903] Chain of Thought Prompting Elicits Reasoning in Large Language Modelsarxiv.org
- [2502.19402] General Reasoning Requires Learning to Reason from the Get-goar5iv.labs.arxiv.org
- Generative AI's Act o1: The Reasoning Era Begins | Sequoia Capitalsequoiacap.com
- Reasoning Models Reason Well, Until They Don'tarxiv.org
- As Rocks May Think | Eric Jangevjang.com
- The State of Reinforcement Learning for LLM Reasoningsebastianraschka.com