HOP to the Next Tasks and Domains for Continual Learning in NLP
Umberto Michieli, Mete Ozay
摘要
Continual Learning (CL) aims to learn a sequence of problems (i.e., tasks and domains) by transferring knowledge acquired on previous problems, whilst avoiding forgetting of past ones. Different from previous approaches which focused on CL for one NLP task or domain in a specific use-case, in this paper, we address a more general CL setting to learn from a sequence of problems in a unique framework. Our method, HOP, permits to hop across tasks and domains by addressing the CL problem along three directions: (i) we employ a set of adapters to generalize a large pre-trained model to unseen problems, (ii) we compute high-order moments over the distribution of embedded representations to distinguish independent and correlated statistics across different tasks and domains, (iii) we process this enriched information with auxiliary heads specialized for each end problem. Extensive experimental campaign on 4 NLP applications, 5 benchmarks and 2 CL setups demonstrates the effectiveness of our HOP.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper15
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati 等NeurIPS 2020 · 被引用 1,494 次
- LAMOL: LAnguage MOdeling for Lifelong Language LearningFan-Keng Sun, Cheng-Hao Ho, Hung-Yi LeeICLR 2020 · 被引用 247 次
- RECALL: Replay-based Continual Learning in Semantic SegmentationAndrea Maracani, Umberto Michieli, Marco Toldo, Pietro ZanuttighICCV 2021 · 被引用 148 次
- Isolation and Impartial Aggregation: A Paradigm of Incremental Learning without InterferenceYabin Wang, Zhiheng Ma, Zhiwu Huang, Yaowei Wang 等AAAI 2023 · 被引用 72 次
- Continual Learning by Using Information of Each Class HolisticallyWenpeng Hu, Qi Qin, Mengyu Wang, Jinwen Ma 等AAAI 2021 · 被引用 64 次
相关 Paper
- Effective Continual Learning for Text Classification with Lightweight SnapshotsJue Wang, Dajie Dong, Lidan Shou, Ke Chen 等AAAI 2023 · 被引用 4 次
- Pretrained Language Model in Continual Learning: A Comparative StudyTongtong Wu, Massimo Caccia, Zhuang Li, Yuan-Fang Li 等ICLR 2022 · 被引用 76 次
- Adapt Before Continual LearningAojun Lu, Tao Feng, Hangjie Yuan, Chunhui Ding 等AAAI 2026
- Enhancing Visual Continual Learning with Language-Guided SupervisionBolin Ni, Hongbo Zhao, Chenghao Zhang, Ke Hu 等CVPR 2024 · 被引用 8 次
- Lifelong Domain Adaptation via Consolidated Internal DistributionMohammad RostamiNeurIPS 2021 · 被引用 72 次
