ReAct: Synergizing Reasoning and Acting in Language Models – Google AI Blog
Posted by Shunyu Yao, Student Researcher, and Yuan Cao, Research Scientist, Google Research, Brain Team Recent advances have expanded the applicability of language models (LM) to downstream tasks. On one hand, existing language models that are properly prompted, via chain-of-thought, demonstrate emergent capabilities that carry out self-conditioned reasoning traces to derive answers from questions, excelling at various arithmetic, commonsense, and symbolic reasoning tasks. However, with chain-of-thought prompting, a model is not grounded in the external world and uses its own internal representations to generate reasoning traces, limiting its ability to reactively explore and reason or update its knowledge. On the other hand, recent work uses pre-trained language models for planning and acting in various interactive environments (e.g., text games, web navigation, embodied tasks, robotics), with a focus on mapping text contexts to text actions via the language model’s internal knowledge
ReAct: Synergizing Reasoning and Acting in Language Models Skip to main content ReAct: Synergizing Reasoning and Acting in Language Models November 8, 2022 Posted by Shunyu Yao, Student Researcher, and Yuan Cao, Research Scientist, Google Research, Brain Team Quick links Share Copy link × --> Recent advances have expanded the applicability of language models (LM) to downstream tasks. On one hand, existing language models that are properly prompted, via chain-of-thought , demonstrate emergent capabilities that carry out self-conditioned reasoning traces to derive answers from questions, excelli
saved by
related reading
- 2210.03629.pdfarxiv.org
- Tracing the thoughts of a large language model \ Anthropicanthropic.com
- DeepSeek-R1arxiv.org
- [2201.11903] Chain of Thought Prompting Elicits Reasoning in Large Language Modelsarxiv.org
- Explore | alphaXivalphaxiv.org
- As Rocks May Think | Eric Jangevjang.com
- Language Models can Solve Computer Tasksarxiv.org
- the-illusion-of-thinking.pdfml-site.cdn-apple.com
- Training Large Language Models to Reason in a Continuous Latent Spacearxiv.org
- [2203.14465] STaR: Bootstrapping Reasoning With Reasoningarxiv.org
- [2412.06769] Training Large Language Models to Reason in a Continuous Latent Spacearxiv.org
- Language Models Perform Reasoning via Chain of Thoughtai.googleblog.com