Continual Sequence Generation with Adaptive Compositional Modules
Yanzhe Zhang, Xuezhi Wang, Diyi Yang
摘要
Continual learning is essential for real-world deployment when there is a need to quickly adapt the model to new tasks without forgetting knowledge of old tasks. Existing work on continual sequence generation either always reuses existing parameters to learn new tasks, which is vulnerable to catastrophic forgetting on dissimilar tasks, or blindly adds new parameters for every new task, which could prevent knowledge sharing between similar tasks. To get the best of both worlds, in this work, we propose continual sequence generation with adaptive compositional modules to adaptively add modules in transformer architectures and compose both old and new modules for new tasks. We also incorporate pseudo experience replay to facilitate knowledge transfer in those shared modules. Experiment results on various sequences of generation tasks show that our framework can adaptively add modules or reuse modules based on task similarity, outperforming state-of-the-art baselines in terms of both performance and parameter efficiency. We make our code public at https://github.com/GT-SALT/ Adaptive-Compositional-Modules .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Mitigating the Alignment Tax of RLHFYong Lin, Hangyu Lin, Wei Xiong, Shizhe Diao 等EMNLP 2024 · 被引用 18 次
- Prompts Can Play Lottery Tickets Well: Achieving Lifelong Information Extraction via Lottery Prompt TuningZujie Liang, Feng Wei, Yin Jie, Yuxi Qian 等ACL 2023 · 被引用 8 次
- Ask and Remember: A Questions-Only Replay Strategy for Continual Visual Question AnsweringImad Eddine Marouf, Enzo Tartaglione, Stéphane Lathuilière, Joost van de WeijerICCV 2025 · 被引用 4 次
- Spurious Forgetting in Continual Learning of Language ModelsJunhao Zheng, Xidi Cai, Shengjie Qiu, Qianli MaICLR 2025
- HiCL: Hippocampal-Inspired Continual LearningKushal Kapoor, Wyatt Mackey, Yiannis Aloimonos, Xiaomin LinAAAI 2026
它引用的顶会 Paper9
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- MixText: Linguistically-Informed Interpolation of Hidden Space for Semi-Supervised Text ClassificationJiaao Chen, Zichao Yang, Diyi YangACL 2020 · 被引用 340 次
- LAMOL: LAnguage MOdeling for Lifelong Language LearningFan-Keng Sun, Cheng-Hao Ho, Hung-Yi LeeICLR 2020 · 被引用 247 次
- Continual Relation Learning via Episodic Memory Activation and ReconsolidationXu Han, Yi Dai, Tianyu Gao, Yankai Lin 等ACL 2020 · 被引用 92 次
- Continual Learning in Task-Oriented Dialogue SystemsAndrea Madotto, Zhaojiang Lin, Zhenpeng Zhou, Seungwhan Moon 等EMNLP 2021 · 被引用 68 次
相关 Paper
- Progressive Prompts: Continual Learning for Language ModelsAnastasia Razdaibiedina, Yuning Mao, Rui Hou, Madian Khabsa 等ICLR 2023 · 被引用 15 次
- Lifelong Sequence Generation with Dynamic Module Expansion and AdaptationChengwei Qin, Chen Chen, Shafiq JotyEMNLP 2023 · 被引用 1 次
- Compositional Language Continual LearningYuanpeng Li, Liang Zhao, Kenneth Church, Mohamed ElhoseinyICLR 2020 · 被引用 40 次
- Efficient Continual Learning with Modular Networks and Task-Driven PriorsTom Veniat, Ludovic Denoyer, Marc'Aurelio RanzatoICLR 2021 · 被引用 110 次
- Self-Expansion of Pre-trained Models with Mixture of Adapters for Continual LearningHuiyi Wang, Haodong Lu, Lina Yao, Dong GongCVPR 2025
