[1606.03126] Key-Value Memory Networks for Directly Reading Documents
Directly reading documents and being able to answer questions from them is an unsolved challenge. To avoid its inherent difficulty, question answering (QA) has been directed towards using Knowledge Bases (KBs) instead, which has proven effective. Unfortunately KBs often suffer from being too restrictive, as the schema cannot support certain types of answers, and too sparse, e.g. Wikipedia contains much more information than Freebase. In this work we introduce a new method, Key-Value Memory Networks, that makes reading documents more viable by utilizing different encodings in the addressing and output stages of the memory read operation. To compare using KBs, information extraction or Wikipedia documents directly in a single framework we construct an analysis tool, WikiMovies, a QA dataset that contains raw text alongside a preprocessed KB, in the domain of movies. Our method reduces the gap between all three settings. It also achieves state-of-the-art results on the existing WikiQA benchmark.
Directly reading documents and being able to answer questions from them is an unsolved challenge. To avoid its inherent difficulty, question answering (QA) has been directed towards using Knowledge Bases (KBs) instead, which has proven effective. Unfortunately KBs often suffer from being too restrictive, as the schema cannot support certain types of answers, and too sparse, e.g. Wikipedia contains much more information than Freebase. In this work we introduce a new method, Key-Value Memory Networks, that makes reading documents more viable by utilizing different encodings in the addressing and
Explore this link on the map →saved by
related reading
- How can we develop transformative tools for thought?numinous.productions
- Key-value memory in the brainarxiv.org
- [2005.11401] Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasksarxiv.org
- Augmenting Long-term Memoryaugmentingcognition.com
- KV Caching Explained: Optimizing Transformer Inference Efficiencyhuggingface.co
- [2605.15156] MeMo: Memory as a Modelarxiv.org
- Augmenting Long-term Memoryaugmentingcognition.com
- GitHub - mem0ai/mem0: Universal memory layer for AI Agents · GitHubgithub.com
- Andrej Karpathy on X: "LLM Knowledge Bases Something I'm finding very useful recently: using LLMs to build personal knowledge bases for various topics of research interest. In this way, a large fraction of my recent token throughput is going less into manipulating code, and more into manipulating" / Xx.com
- TurboQuant: Redefining AI efficiency with extreme compressionresearch.google
- GitHub - Future-House/paper-qa: High accuracy RAG for answering questions from scientific documents with citations · GitHubgithub.com
- Memex - Save, summarize and reuse what you read online.memex.garden