Mention Memory: incorporating textual knowledge into Transformers through entity mention attention
Michiel de Jong, Yury Zemlyanskiy, Nicholas FitzGerald, Fei Sha, William W. Cohen
Abstract
Natural language understanding tasks such as open-domain question answering often require retrieving and assimilating factual information from multiple sources. We propose to address this problem by integrating a semi-parametric representation of a large text corpus into a Transformer model as a source of factual knowledge. Specifically, our method represents knowledge with `mention memory', a table of dense vector representations of every entity mention in a corpus. The proposed model - TOME - is a Transformer that accesses the information through internal memory layers in which each entity mention in the input passage attends to the mention memory. This approach enables synthesis of and reasoning over many disparate sources of information within a single Transformer model. In experiments using a memory of 150 million Wikipedia mentions, TOME achieves strong performance on several open-domain knowledge-intensive tasks, including the claim verification benchmarks HoVer and FEVER and several entity-based QA benchmarks. We also show that the model learns to attend to informative mentions without any direct supervision. Finally we demonstrate that the model can generalize to new unseen entities by updating the memory without retraining.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 27351645-ed0e-443b-b15a-353b3e5d6031Cited by top-tier papers14
- MuRAG: Multimodal Retrieval-Augmented Generator for Open Question Answering over Images and TextWenhu Chen, Hexiang Hu, Xi Chen, Pat Verga et al.EMNLP 2022 · 89 citations
- Training Language Models with Memory AugmentationZexuan Zhong, Tao Lei, Danqi ChenEMNLP 2022 · 52 citations
- Fine-tuning Image Transformers using Learnable MemoryMark Sandler, Andrey Zhmoginov, Max Vladymyrov, Andrew JacksonCVPR 2022 · 51 citations
- Decoupled Context Processing for Context Augmented Language ModelingZonglin Li, Ruiqi Guo, Sanjiv KumarNeurIPS 2022 · 31 citations
- Pre-computed memory or on-the-fly encoding? A hybrid approach to retrieval augmentation makes the most of your computeMichiel de Jong, Yury Zemlyanskiy, Nicholas FitzGerald, Joshua Ainslie et al.ICML 2023 · 20 citations
Builds on9
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni et al.NeurIPS 2020 · 19,162 citations
- Retrieval Augmented Language Model Pre-TrainingKelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat et al.ICML 2020 · 2,937 citations
- Accelerating Large-Scale Inference with Anisotropic Vector QuantizationRuiqi Guo, Philip Sun, Erik Lindgren, Quan Geng et al.ICML 2020 · 539 citations
- Pre-training via ParaphrasingMike Lewis, Marjan Ghazvininejad, Gargi Ghosh, Armen Aghajanyan et al.NeurIPS 2020 · 165 citations
- Dense Passage Retrieval for Open-Domain Question AnsweringVladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis et al.EMNLP 2020 · 142 citations
Related papers
- An Efficient Memory-Augmented Transformer for Knowledge-Intensive NLP TasksYuxiang Wu, Yu Zhao, Baotian Hu, Pasquale Minervini et al.EMNLP 2022 · 29 citations
- A Unified Encoder-Decoder Framework with Entity MemoryZhihan Zhang, Wenhao Yu, Chenguang Zhu, Meng JiangEMNLP 2022 · 9 citations
- Entities as Experts: Sparse Memory Access with Entity SupervisionThibault Févry, Livio Baldini Soares, Nicholas FitzGerald, Eunsol Choi et al.EMNLP 2020 · 39 citations
- TANDA: Transfer and Adapt Pre-Trained Transformer Models for Answer Sentence SelectionSiddhant Garg, Thuy Vu, Alessandro MoschittiAAAI 2020 · 229 citations
- Contextualize Knowledge Bases with Transformer for End-to-end Task-Oriented Dialogue SystemsYanjie Gou, Yinjie Lei, Lingqiao Liu, Yong Dai et al.EMNLP 2021 · 11 citations
