Preserving Generalization of Language models in Few-shot Continual Relation Extraction
Quyen Tran, Nguyen Xuan Thanh, Nguyen Hoang Anh, Nam Le Hai, Trung Le, Linh Van Ngo, Thien Huu Nguyen
Abstract
Few-shot Continual Relations Extraction (FCRE) is an emerging and dynamic area of study where models can sequentially integrate knowledge from new relations with limited labeled data while circumventing catastrophic forgetting and preserving prior knowledge from pre-trained backbones. In this work, we introduce a novel method that leverages often-discarded language model heads. By employing these components via a mutual information maximization strategy, our approach helps maintain prior knowledge from the pre-trained backbone and strategically aligns the primary classification head, thereby enhancing model performance. Furthermore, we explore the potential of Large Language Models (LLMs), renowned for their wealth of knowledge, in addressing FCRE challenges. Our comprehensive experimental results underscore the efficacy of the proposed method and offer valuable insights for future work.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d21126c7-b9b1-4c53-92dc-7f2a9613d5fcCited by top-tier papers5
- Adaptive Prompting for Continual Relation Extraction: A Within-Task Variance PerspectiveMinh Le, Tien Ngoc Luu, An Nguyen The, Thanh-Thien Le et al.AAAI 2025 · 12 citations
- Few-Shot, No Problem: Descriptive Continual Relation ExtractionNguyen Xuan Thanh, Anh Duc Le, Quyen Tran, Thanh-Thien Le et al.AAAI 2025 · 6 citations
- An Optimal Transport-driven Approach for Cultivating Latent Space in Online Incremental LearningQuyen Tran, Hai Nguyen, Minh Quan Dao, Hoang Phan et al.CVPR 2026
- TALAS: Teacher-Anchored Layer Alignment with Adaptive Sharpness-Aware Minimization for Embedding DistillationQuoc Phong Dao, Hoang Son Nguyen, Pham Khanh Chi, Linh Ngo Van et al.ACL 2026
- Mitigating Non-Representative Prototypes and Representation Bias in Few-Shot Continual Relation ExtractionThanh Duc Pham, Nam Le Hai, Linh Ngo Van, Nguyen Thi Ngoc Diep et al.ACL 2025
Builds on6
- Calibrate Before Use: Improving Few-shot Performance of Language ModelsZihao Zhao, Eric Wallace, Shi Feng, Dan Klein et al.ICML 2021 · 1,843 citations
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati et al.NeurIPS 2020 · 1,494 citations
- Online Continual Learning through Mutual Information MaximizationYiduo Guo, Bing Liu, Dongyan ZhaoICML 2022 · 139 citations
- Continual Relation Learning via Episodic Memory Activation and ReconsolidationXu Han, Yi Dai, Tianyu Gao, Yankai Lin et al.ACL 2020 · 92 citations
- Continual Few-shot Relation Learning via Embedding Space Regularization and Data AugmentationChengwei Qin, Shafiq R. JotyACL 2022 · 49 citations
Related papers
- Consistent Prototype Learning for Few-Shot Continual Relation ExtractionXiudi Chen, Hui Wu, Xiaodong ShiACL 2023 · 17 citations
- RECALL: REpresentation-aligned Catastrophic-forgetting ALLeviation via Hierarchical Model MergingBowen Wang, Haiyuan Wan, Liwen Shi, Chen Yang et al.EMNLP 2025
- Few-shot Continual Infomax LearningZiqi Gu, Chunyan Xu, Jian Yang, Zhen CuiICCV 2023 · 18 citations
- SEEKR: Selective Attention-Guided Knowledge Retention for Continual Learning of Large Language ModelsJinghan He, Haiyun Guo, Kuan Zhu, Zihan Zhao et al.EMNLP 2024 · 4 citations
- Learning Robust Representations for Continual Relation Extraction via Adversarial Class AugmentationPeiyi Wang, Yifan Song, Tianyu Liu, Binghuai Lin et al.EMNLP 2022 · 23 citations
