Continual Learning in Token Space | Letta
Create, deploy, and manage your agents at scale with Letta Cloud. Build production applications backed by agent microservices with REST APIs. Letta adds memory to your LLM services to give them advanced reasoning capabilities and transparent long-term memory (powered by MemGPT). The biggest gap between AI agents and human intelligence is the ability to learn. Humans continually learn and improve over time, acquire new skills, update their beliefs based on new facts, and modify their behavior to correct for past mistakes. In contrast, most AI agents have an incredible amount of world knowledge, but do not meaningfully get better over time. How do we create AI agents that can continually learn? Traditionally, the concept of “continual learning” for neural networks has been synonymous with weight updates, under the assumption that all learning happens in a connectionist way. The central research questions have focused on catastrophic forgetting (new weight updates causing accidental knowl
BRIEF The continual learning problem in LLM agents is best viewed through the lens of learning in token space: updates to learned context, not weights, should be the primary mechanism for LLM agents to learn from experience. The biggest gap between AI agents and human intelligence is the ability to learn. Humans continually learn and improve over time, acquire new skills, update their beliefs based on new facts, and modify their behavior to correct for past mistakes. In contrast, most AI agents have an incredible amount of world knowledge, but do not meaningfully get better over time. How do w
Explore this link on the map →saved by
related reading
- Why We Need Continual Learning | Andreessen Horowitza16z.com
- The Continual Learning Problemjessylin.com
- Andrej Karpathy — AGI is still a decade awaydwarkesh.com
- Effective context engineering for AI agents \ Anthropicanthropic.com
- Context Engineering for AI Agents: Lessons from Building Manusmanus.im
- NL.pdfabehrouz.github.io
- You can’t imitation-learn how to continual-learn — LessWronglesswrong.com
- What are the real problems of continual learning?infinitefaculty.substack.com
- Reimagining LLM Memory: Using Context as Training Data Unlocks Models That Learn at Test-Time | NVIDIA Technical Blogdeveloper.nvidia.com
- What's so hard about continuous learning?seangoedecke.com
- Introducing Context Repositories: Git-based Memory for Coding Agents | Lettaletta.com
- augustus odena on X: "I have a bunch of thoughts about continual learning and nothing to do with them (I'm working on something else) so I figured I'd just turn them into a post: First: I think people use "continual learning" to point at a cluster of issues that are related but distinct. I'll list" / Xx.com