flâneur — a map of the web's best reading

Language Models Perform Reasoning via Chain of Thought – Google AI Blog

ai.googleblog.com · 1,127 words · saved by 1 readers

In recent years, scaling up the size of language models has been shown to be a reliable way to improve performance on a range of natural language processing (NLP) tasks. Today’s language models at the scale of 100B or more parameters achieve strong performance on tasks like sentiment analysis and machine translation, even with little or no training examples. Even the largest language models, however, can still struggle with certain multi-step reasoning tasks, such as math word problems and commonsense reasoning. How might we enable language models to perform such reasoning tasks? In “Chain of Thought Prompting Elicits Reasoning in Large Language Models,” we explore a prompting method for improving the reasoning abilities of language models. Called chain of thought prompting, this method enables models to decompose multi-step problems into intermediate steps. With chain of thought prompting, language models of sufficient scale (~100B parameters) can solve complex reasoning problems that

Language Models Perform Reasoning via Chain of Thought Skip to main content Language Models Perform Reasoning via Chain of Thought May 11, 2022 Posted by Jason Wei and Denny Zhou, Research Scientists, Google Research, Brain team Quick links Share Copy link × In recent years, scaling up the size of language models has been shown to be a reliable way to improve performance on a range of natural language processing (NLP) tasks. Today’s language models at the scale of 100B or more parameters achieve strong performance on tasks like sentiment analysis and machine translation, even with little or no

Explore this link on the map →

related reading