Progressive Memory Banks for Incremental Domain Adaptation
Nabiha Asghar, Lili Mou, Kira A. Selby, Kevin D. Pantasdo, Pascal Poupart, Xin Jiang
摘要
This paper addresses the problem of incremental domain adaptation (IDA) in natural language processing (NLP). We assume each domain comes one after another, and that we could only access data in the current domain. The goal of IDA is to build a unified model performing well on all the domains that we have encountered. We adopt the recurrent neural network (RNN) widely used in NLP, but augment it with a directly parameterized memory bank, which is retrieved by an attention mechanism at each step of RNN transition. The memory bank provides a natural way of IDA: when adapting our model to a new domain, we progressively add new slots to the memory bank, which increases the number of parameters, and thus the model capacity. We learn the new memory slots and fine-tune existing parameters by back-propagation. Experimental results show that our approach achieves significantly better performance than fine-tuning alone. Compared with expanding hidden states, our approach is more robust for old domains, shown by both empirical and theoretical results. Our model also outperforms previous work of IDA including elastic weight consolidation and progressive neural networks in the experiments. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Continual learning in recurrent neural networksBenjamin Ehret, Christian Henning, Maria R. Cervera, Alexander Meulemans 等ICLR 2021 · 被引用 4,433 次
- Memory-Augmented Non-Local Attention for Video Super-ResolutionJiyang Yu, Jingen Liu, Liefeng Bo, Tao MeiCVPR 2022 · 被引用 47 次
- TokMem: One-Token Procedural Memory for Large Language ModelsZijun Wu, Yongchang Hao, Lili MouICLR 2026 · 被引用 4 次
- HOP to the Next Tasks and Domains for Continual Learning in NLPUmberto Michieli, Mete OzayAAAI 2024 · 被引用 3 次
- Memory-Based Invariance Learning for Out-of-Domain Text ClassificationChen Jia, Yue ZhangEMNLP 2023 · 被引用 1 次
相关 Paper
- A Unified Approach to Domain Incremental Learning with Memory: Theory and AlgorithmHaizhou Shi, Hao WangNeurIPS 2023 · 被引用 60 次
- Continual Pre-training of Language ModelsZixuan Ke, Yijia Shao, Haowei Lin, Tatsuya Konishi 等ICLR 2023 · 被引用 15 次
- A Simple Yet Effective Subsequence-Enhanced Approach for Cross-Domain NERJinpeng Hu, Dandan Guo, Yang Liu, Zhuo Li 等AAAI 2023 · 被引用 11 次
- Three Heads Are Better than One: Improving Cross-Domain NER with Progressive Decomposed NetworkXuming Hu, Zhaochen Hong, Yong Jiang, Zhichao Lin 等AAAI 2024 · 被引用 1 次
- Non-exemplar Domain Incremental Object Detection via Learning Domain BiasXiang Song, Yuhang He, Songlin Dong, Yihong GongAAAI 2024 · 被引用 14 次
