Learning to Rehearse in Long Sequence Memorization
Zhu Zhang, Chang Zhou, Jianxin Ma, Zhijie Lin, Jingren Zhou, Hongxia Yang, Zhou Zhao
摘要
Existing reasoning tasks often have an important assumption that the input contents can be always accessed while reasoning, requiring unlimited storage resources and suffering from severe time delay on long sequences. To achieve efficient reasoning on long sequences with limited storage resources, memory augmented neural networks introduce a human-like write-read memory to compress and memorize the long input sequence in one pass, trying to answer subsequent queries only based on the memory. But they have two serious drawbacks: 1) they continually update the memory from current information and inevitably forget the early contents; 2) they do not distinguish what information is important and treat all contents equally. In this paper, we propose the Rehearsal Memory (RM) to enhance long-sequence memorization by self-supervised rehearsal with a history sampler. To alleviate the gradual forgetting of early information, we design self-supervised rehearsal training with recollection and familiarity tasks. Further, we design a history sampler to select informative fragments for rehearsal training, making the memory focus on the crucial information. We evaluate the performance of our rehearsal memory by the synthetic bAbI task and several downstream tasks, including text/video question answering and recommendation on long sequences.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Why Do We Click: Visual Impression-aware News RecommendationJiahao Xun, Shengyu Zhang, Zhou Zhao, Jieming Zhu 等ACM MM 2021 · 被引用 28 次
- Track-On: Transformer-based Online Point Tracking with MemoryGörkay Aydemir, Xiongyi Cai, Weidi Xie, Fatma GüneyICLR 2025
- Grounded Question-Answering in Long Egocentric VideosShangzhe Di, Weidi XieCVPR 2024
它引用的顶会 Paper7
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Compressive Transformers for Long-Range Sequence ModellingJack W. Rae, Anna Potapenko, Siddhant M. Jayakumar, Chloe Hillier 等ICLR 2020 · 被引用 833 次
- CauseRec: Counterfactual User Sequence Synthesis for Sequential RecommendationShengyu Zhang, Dong Yao, Zhou Zhao, Tat-Seng Chua 等SIGIR 2021 · 被引用 118 次
- Self-Attentive Associative MemoryHung Le, Truyen Tran, Svetha VenkateshICML 2020 · 被引用 61 次
- Neural Stored-program MemoryHung Le, Truyen Tran, Svetha VenkateshICLR 2020 · 被引用 38 次
相关 Paper
- MEMO: A Deep Network for Flexible Combination of Episodic MemoriesAndrea Banino, Adrià Puigdomènech Badia, Raphael Köster, Martin J. Chadwick 等ICLR 2020 · 被引用 41 次
- Recurrent Memory TransformerAydar Bulatov, Yuri Kuratov, Mikhail BurtsevNeurIPS 2022 · 被引用 252 次
- GradMem: Learning to Write Context into Memory with Test-Time Gradient DescentYuri Kuratov, Matvey Kairov, Aydar Bulatov, Ivan Rodkin 等ICML 2026 · 被引用 3 次
- Enlarging the Long-time Dependencies via RL-based Memory Network in Movie Affective AnalysisJie Zhang, Yin Zhao, Kai QianACM MM 2022 · 被引用 4 次
- Dynamic Long Context Reasoning over Compressed Memory via End-to-End Reinforcement LearningZhuoen Chen, Dongfang Li, Meishan Zhang, Baotian Hu 等ACL 2026 · 被引用 2 次
