SKD-NER: Continual Named Entity Recognition via Span-based Knowledge Distillation with Reinforcement Learning
Yi Chen, Liang He
摘要
Continual learning for named entity recognition (CL-NER) aims to enable models to continuously learn new entity types while retaining the ability to recognize previously learned ones. However, the current strategies fall short of effectively addressing the catastrophic forgetting of previously learned entity types. To tackle this issue, we propose the SKD-NER model, an efficient continual learning NER model based on the span-based approach, which innovatively incorporates reinforcement learning strategies to enhance the model's ability against catastrophic forgetting. Specifically, we leverage knowledge distillation (KD) to retain memory and employ reinforcement learning strategies during the KD process to optimize the soft labeling and distillation losses generated by the teacher model to effectively prevent catastrophic forgetting during continual learning. This approach effectively prevents or mitigates catastrophic forgetting during continuous learning, allowing the model to retain previously learned knowledge while acquiring new knowledge. Our experiments on two benchmark datasets demonstrate that our model significantly improves the performance of the CL-NER task, outperforming state-of-the-art methods. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- Continual Learning for Named Entity RecognitionNatawut Monaikul, Giuseppe Castellucci, Simone Filice, Oleg RokhlenkoAAAI 2021 · 被引用 84 次
- Neural Topic Modeling with Continual Lifelong LearningPankaj Gupta, Yatin Chaudhary, Thomas A. Runkler, Hinrich SchützeICML 2020 · 被引用 55 次
- Is Reinforcement Learning (Not) for Natural Language Processing: Benchmarks, Baselines, and Building Blocks for Natural Language Policy OptimizationRajkumar Ramamurthy, Prithviraj Ammanabrolu, Kianté Brantley, Jack Hessel 等ICLR 2023 · 被引用 54 次
- Distilling Causal Effect from Miscellaneous Other-Class for Continual Named Entity RecognitionJunhao Zheng, Zhanxian Liang, Haibin Chen, Qianli MaEMNLP 2022 · 被引用 20 次
- A Neural Span-Based Continual Named Entity Recognition ModelYunan Zhang, Qingcai ChenAAAI 2023 · 被引用 16 次
相关 Paper
- Continual Named Entity Recognition without Catastrophic ForgettingDuzhen Zhang, Wei Cong, Jiahua Dong, Yahan Yu 等EMNLP 2023 · 被引用 16 次
- Unify Named Entity Recognition Scenarios via Contrastive Real-Time Updating PrototypeYanhe Liu, Peng Wang, Wenjun Ke, Guozheng Li 等AAAI 2024 · 被引用 7 次
- Prototypical Replay with Old-class Focusing Knowledge Distillation for Incremental Named Entity RecognitionZesheng Liu, Qiannan Zhu, Cuiping Li, Hong ChenAAAI 2025
- SEEKR: Selective Attention-Guided Knowledge Retention for Continual Learning of Large Language ModelsJinghan He, Haiyun Guo, Kuan Zhu, Zihan Zhao 等EMNLP 2024 · 被引用 4 次
- Few-Shot Class-Incremental Learning for Named Entity RecognitionRui Wang, Tong Yu, Handong Zhao, Sungchul Kim 等ACL 2022 · 被引用 26 次
