Learning to Learn and Remember Super Long Multi-Domain Task Sequence
Zhenyi Wang, Li Shen, Tiehang Duan, Donglin Zhan, Le Fang, Mingchen Gao
摘要
Catastrophic forgetting (CF) frequently occurs when learning with non-stationary data distribution. The CF issue remains nearly unexplored and is more challenging when meta-learning on a sequence of domains (datasets), called sequential domain meta-learning (SDML). In this work, we propose a simple yet effective learning to learn approach, i.e., meta optimizer, to mitigate the CF problem in SDML. We first apply the proposed meta optimizer to the simplified setting of SDML, domain-aware meta-learning, where the domain labels and boundaries are known during the learning process. We propose dynamically freezing the network and incorporating it with the proposed meta optimizer by considering the domain nature during meta training. In addition, we extend the meta optimizer to the more general setting of SDML, domain-agnostic meta-learning, where domain labels and boundaries are unknown during the learning process. We propose a domain shift detection technique to capture latent domain change and equip the meta optimizer with it to work in this setting. The proposed meta optimizer is versatile and can be easily integrated with several existing meta-learning algorithms. Finally, we construct a challenging and large-scale benchmark consisting of 10 heterogeneous domains with a super long task sequence consisting of 100K tasks. We perform extensive experiments on the proposed benchmark for both settings and demonstrate the effectiveness of our proposed method, outperforming current strong baselines by a large margin.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Improving Task-free Continual Learning by Distributionally Robust Memory EvolutionZhenyi Wang, Li Shen, Le Fang, Qiuling Suo 等ICML 2022 · 被引用 52 次
- Probabilistic Bilevel Coreset SelectionXiao Zhou, Renjie Pi, Weizhong Zhang, Yong Lin 等ICML 2022 · 被引用 39 次
- On the Stability-Plasticity Dilemma in Continual Meta-Learning: Theory and AlgorithmQi Chen, Changjian Shui, Ligong Han, Mario MarchandNeurIPS 2023 · 被引用 32 次
- Data Augmented Flatness-aware Gradient Projection for Continual LearningEnneng Yang, Li Shen, Zhenyi Wang, Shiwei Liu 等ICCV 2023 · 被引用 28 次
- Learning to Learn from APIs: Black-Box Data-Free Meta-LearningZixuan Hu, Li Shen, Zhenyi Wang, Baoyuan Wu 等ICML 2023 · 被引用 18 次
它引用的顶会 Paper24
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine 等NeurIPS 2020 · 被引用 2,261 次
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati 等NeurIPS 2020 · 被引用 1,494 次
- Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAMLAniruddh Raghu, Maithra Raghu, Samy Bengio, Oriol VinyalsICLR 2020 · 被引用 736 次
- Coresets for Data-efficient Training of Machine Learning ModelsBaharan Mirzasoleiman, Jeff A. Bilmes, Jure LeskovecICML 2020 · 被引用 494 次
- Cross-Domain Few-Shot Classification via Learned Feature-Wise TransformationHung-Yu Tseng, Hsin-Ying Lee, Jia-Bin Huang, Ming-Hsuan YangICLR 2020 · 被引用 467 次
相关 Paper
- Learning to Adapt to Evolving DomainsHong Liu, Mingsheng Long, Jianmin Wang, Yu WangNeurIPS 2020 · 被引用 63 次
- Meta Learning on a Sequence of Imbalanced Domains with Difficulty AwarenessZhenyi Wang, Tiehang Duan, Le Fang, Qiuling Suo 等ICCV 2021 · 被引用 21 次
- Learning to Continually Learn with the Bayesian PrincipleSoochan Lee, Hyeonseong Jeon, Jaehyeon Son, Gunhee KimICML 2024 · 被引用 11 次
- Continual Adaptation of Visual Representations via Domain Randomization and Meta-LearningRiccardo Volpi, Diane Larlus, Grégory RogezCVPR 2021
- ACE: Adapting to Changing Environments for Semantic SegmentationZuxuan Wu, Xin Wang, Joseph Gonzalez, Tom Goldstein 等ICCV 2019 · 被引用 109 次
