Do Large Language Models (LLMs) reason? | Shaped Blog
In recent years, Large Language Models (LLMs) have revolutionized the field of Natural Language Processing (NLP), enabling significant advancements in language understanding, text generation, and more. With the help of memorization and compositionality capabilities, LLMs can perform various tasks like never before. They're already at the core of several products used by millions of people, such as Google's search engine, Github's Copilot, and OpenAI's ChatGPT! Title image from Xavier Amatriain (2023) Despite their groundbreaking capabilities, people argue whether these language models are still missing fundamental parts that make up general intelligence. The question is whether these language models, when the parameters are scaled up, could match human-level intelligence. Or are these language models just understanding the statistics of language, so that they can pattern-match output well enough to mimic understanding? In the recent paper: “Augmented Language Models: a Survey” from Met
Do Large Language Models (LLMs) reason? In recent years, Large Language Models (LLMs) have revolutionized the field of Natural Language Processing (NLP), enabling significant advancements in language understanding, text generation, and more. With the help of memorization and compositionality capabilities, LLMs can perform various tasks like never before. They're already at the core of several products used by millions of people, such as Google's search engine, Github's Copilot, and OpenAI's ChatGPT! Feb 21, 2023 | 10 min read by Nina Shenker Tauris Title image from Xavier Amatriain (2023) Desp
Explore this link on the map →related reading
- [2201.11903] Chain of Thought Prompting Elicits Reasoning in Large Language Modelsarxiv.org
- DeepSeek-R1arxiv.org
- Tracing the thoughts of a large language model \ Anthropicanthropic.com
- the-illusion-of-thinking.pdfml-site.cdn-apple.com
- The Role of Deductive and Inductive Reasoning in Large Language Models - ACL Anthologyaclanthology.org
- [2205.11916] Large Language Models are Zero-Shot Reasonersarxiv.org
- [2502.19402] General Reasoning Requires Learning to Reason from the Get-goar5iv.labs.arxiv.org
- Unlocking the Working Memory of Large Language Models for Latent Reasoningarxiv.org
- [2501.12948] DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learningarxiv.org
- [2201.11903] Chain-of-Thought Prompting Elicits Reasoning in Large Language Modelsarxiv.org
- Training Large Language Models to Reason in a Continuous Latent Spacearxiv.org
- Language Models can Solve Computer Tasksarxiv.org