[2309.02427] Cognitive Architectures for Language Agents
Abstract:Recent efforts have augmented large language models (LLMs) with external resources (e.g., the Internet) or internal control flows (e.g., prompt chaining) for tasks requiring grounding or reasoning, leading to a new class of language agents. While these agents have achieved substantial empirical success, we lack a systematic framework to organize existing agents and plan future developments. In this paper, we draw on the rich history of cognitive science and symbolic artificial intelligence to propose Cognitive Architectures for Language Agents (CoALA). CoALA describes a language agent with modular memory components, a structured action space to interact with internal memory and external environments, and a generalized decision-making process to choose actions. We use CoALA to retrospectively survey and organize a large body of recent work, and prospectively identify actionable directions towards more capable agents. Taken together, CoALA contextualizes today's language agents within the broader history of AI and outlines a path towards language-based general intelligence.
Abstract:Recent efforts have augmented large language models (LLMs) with external resources (e.g., the Internet) or internal control flows (e.g., prompt chaining) for tasks requiring grounding or reasoning, leading to a new class of language agents. While these agents have achieved substantial empirical success, we lack a systematic framework to organize existing agents and plan future developments. In this paper, we draw on the rich history of cognitive science and symbolic artificial intelligence to propose Cognitive Architectures for Language Agents (CoALA). CoALA describes a language agent
Explore this link on the map →related reading
- LLM Powered Autonomous Agents | Lil'Loglilianweng.github.io
- Tracing the thoughts of a large language model \ Anthropicanthropic.com
- Building Effective AI Agents \ Anthropicanthropic.com
- Building Effective AI Agents \ Anthropicanthropic.com
- Language Models can Solve Computer Tasksarxiv.org
- Don’t Build Multi-Agents | Cognitioncognition.ai
- Explore | alphaXivalphaxiv.org
- Language as a Cognitive Tool: Dall-E, Humans and Vygotskian RL Agents – Developmental Systems, a Blog of the Flowers Labdevelopmentalsystems.org
- Language Models in Plato's Cave - by Sergey Levinesergeylevine.substack.com
- Externalized reasoning oversight: a research direction for language model alignment — AI Alignment Forumalignmentforum.org
- ReAct: Synergizing Reasoning and Acting in Language Modelsai.googleblog.com
- Large Language Model: world models or surface statistics?thegradient.pub