Language Models Perform Reasoning via Chain of Thought – Google AI Blog
In recent years, scaling up the size of language models has been shown to be a reliable way to improve performance on a range of natural language processing (NLP) tasks. Today’s language models at the scale of 100B or more parameters achieve strong performance on tasks like sentiment analysis and machine translation, even with little or no training examples. Even the largest language models, however, can still struggle with certain multi-step reasoning tasks, such as math word problems and commonsense reasoning. How might we enable language models to perform such reasoning tasks? In “Chain of Thought Prompting Elicits Reasoning in Large Language Models,” we explore a prompting method for improving the reasoning abilities of language models. Called chain of thought prompting, this method enables models to decompose multi-step problems into intermediate steps. With chain of thought prompting, language models of sufficient scale (~100B parameters) can solve complex reasoning problems that
Language Models Perform Reasoning via Chain of Thought Skip to main content Language Models Perform Reasoning via Chain of Thought May 11, 2022 Posted by Jason Wei and Denny Zhou, Research Scientists, Google Research, Brain team Quick links Share Copy link × In recent years, scaling up the size of language models has been shown to be a reliable way to improve performance on a range of natural language processing (NLP) tasks. Today’s language models at the scale of 100B or more parameters achieve strong performance on tasks like sentiment analysis and machine translation, even with little or no
Explore this link on the map →related reading
- [2201.11903] Chain of Thought Prompting Elicits Reasoning in Large Language Modelsarxiv.org
- [2201.11903] Chain-of-Thought Prompting Elicits Reasoning in Large Language Modelsarxiv.org
- Pathways Language Model (PaLM): Scaling to 540 Billion Parameters for Breakthrouai.googleblog.com
- Tracing the thoughts of a large language model \ Anthropicanthropic.com
- the-illusion-of-thinking.pdfml-site.cdn-apple.com
- DeepSeek-R1arxiv.org
- Chain-of-Thought Promptinglearnprompting.org
- Explore | alphaXivalphaxiv.org
- Reasoning as Trajectoriesslhleosun.github.io
- Language Models can Solve Computer Tasksarxiv.org
- Do Large Language Models (LLMs) reason? | Shapedshaped.ai
- The Unintelligibility is Ours: Notes on Chain of Thought1a3orn.com