SKD-NER: Continual Named Entity Recognition via Span-based Knowledge Distillation with Reinforcement Learning
Yi Chen, Liang He
Abstract
Continual learning for named entity recognition (CL-NER) aims to enable models to continuously learn new entity types while retaining the ability to recognize previously learned ones. However, the current strategies fall short of effectively addressing the catastrophic forgetting of previously learned entity types. To tackle this issue, we propose the SKD-NER model, an efficient continual learning NER model based on the span-based approach, which innovatively incorporates reinforcement learning strategies to enhance the model's ability against catastrophic forgetting. Specifically, we leverage knowledge distillation (KD) to retain memory and employ reinforcement learning strategies during the KD process to optimize the soft labeling and distillation losses generated by the teacher model to effectively prevent catastrophic forgetting during continual learning. This approach effectively prevents or mitigates catastrophic forgetting during continuous learning, allowing the model to retain previously learned knowledge while acquiring new knowledge. Our experiments on two benchmark datasets demonstrate that our model significantly improves the performance of the CL-NER task, outperforming state-of-the-art methods. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8bdd72e3-4edd-4dee-9ad2-1f0fac600c3fBuilds on9
- Continual Learning for Named Entity RecognitionNatawut Monaikul, Giuseppe Castellucci, Simone Filice, Oleg RokhlenkoAAAI 2021 · 84 citations
- Neural Topic Modeling with Continual Lifelong LearningPankaj Gupta, Yatin Chaudhary, Thomas A. Runkler, Hinrich SchützeICML 2020 · 55 citations
- Is Reinforcement Learning (Not) for Natural Language Processing: Benchmarks, Baselines, and Building Blocks for Natural Language Policy OptimizationRajkumar Ramamurthy, Prithviraj Ammanabrolu, Kianté Brantley, Jack Hessel et al.ICLR 2023 · 54 citations
- Distilling Causal Effect from Miscellaneous Other-Class for Continual Named Entity RecognitionJunhao Zheng, Zhanxian Liang, Haibin Chen, Qianli MaEMNLP 2022 · 20 citations
- A Neural Span-Based Continual Named Entity Recognition ModelYunan Zhang, Qingcai ChenAAAI 2023 · 16 citations
Related papers
- Continual Named Entity Recognition without Catastrophic ForgettingDuzhen Zhang, Wei Cong, Jiahua Dong, Yahan Yu et al.EMNLP 2023 · 16 citations
- Unify Named Entity Recognition Scenarios via Contrastive Real-Time Updating PrototypeYanhe Liu, Peng Wang, Wenjun Ke, Guozheng Li et al.AAAI 2024 · 7 citations
- Prototypical Replay with Old-class Focusing Knowledge Distillation for Incremental Named Entity RecognitionZesheng Liu, Qiannan Zhu, Cuiping Li, Hong ChenAAAI 2025
- SEEKR: Selective Attention-Guided Knowledge Retention for Continual Learning of Large Language ModelsJinghan He, Haiyun Guo, Kuan Zhu, Zihan Zhao et al.EMNLP 2024 · 4 citations
- Few-Shot Class-Incremental Learning for Named Entity RecognitionRui Wang, Tong Yu, Handong Zhao, Sungchul Kim et al.ACL 2022 · 26 citations
