Memory Machines — can LLMs make flashcards that last?
memory-machines.com · 549 words · saved by 1 readers
We benchmark 16 frontier LLMs on AI flashcards from reading highlights. Even GPT-5.2 produces unusable spaced-repetition prompts 36% of the time.
Read the report The full research writeup — evaluation, training, grounding, and arena results Memory systems make memory a choice—but only if you write practice prompts that effectively reinforce those ideas. That often demands more effort or skill than users can muster—and prompts can’t easily evolve or deepen over time. Could we make memory as effortless as using a highlighter? We explored whether LLMs could convert casual highlights into useful memory prompts. We found that models can usually identify the intent of highlights, but struggle to generate prompts that will survive…
saved by
related reading
- How can we develop transformative tools for thought?numinous.productions
- What We’ve Learned From A Year of Building with LLMs – Applied LLMsapplied-llms.org
- LLM Daydreaming · Gwern.netgwern.net
- MemPrompt: Memory-assisted Prompt Editing with User Feedbackmemprompt.com
- How to write good prompts: using spaced repetition to create understandingandymatuschak.org
- Everything I'll forget about prompting LLMsolickel.com
- Prompts for Work & Play: Launching the Wolfram Prompt Repository-Stephen Wolfram Writingswritings.stephenwolfram.com
- Patterns for Building LLM-based Systems & Productseugeneyan.com
- Building a better memory systemmichaelnotebook.com
- LLM Powered Autonomous Agents | Lil'Loglilianweng.github.io
- GitHub - brexhq/prompt-engineering: Tips and tricks for working with Large Language Models like OpenAI's GPT-4.github.com
- Prompting Fundamentals and How to Apply them Effectivelyeugeneyan.com