Distilling Causal Effect from Miscellaneous Other-Class for Continual Named Entity Recognition
Junhao Zheng, Zhanxian Liang, Haibin Chen, Qianli Ma
摘要
Continual Learning for Named Entity Recognition (CL-NER) aims to learn a growing number of entity types over time from a stream of data. However, simply learning Other-Class in the same way as new entity types amplifies the catastrophic forgetting and leads to a substantial performance drop. The main cause behind this is that Other-Class samples usually contain old entity types, and the old knowledge in these Other-Class samples is not preserved properly. Thanks to the causal inference, we identify that the forgetting is caused by the missing causal effect from the old data. To this end, we propose a unified causal framework to retrieve the causality from both new entity types and Other-Class. Furthermore, we apply curriculum learning to mitigate the impact of label noise and introduce a self-adaptive weight for balancing the causal effects between new entity types and Other-Class. Experimental results on three benchmark datasets show that our method outperforms the state-of-theart method by a large margin. Moreover, our method can be combined with the existing stateof-the-art methods to improve the performance in CL-NER. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Continual Named Entity Recognition without Catastrophic ForgettingDuzhen Zhang, Wei Cong, Jiahua Dong, Yahan Yu 等EMNLP 2023 · 被引用 16 次
- SKD-NER: Continual Named Entity Recognition via Span-based Knowledge Distillation with Reinforcement LearningYi Chen, Liang HeEMNLP 2023 · 被引用 10 次
- COLA: Contextualized Commonsense Causal Reasoning from the Causal Inference PerspectiveZhaowei Wang, Quyet V. Do, Hongming Zhang, Jiayao Zhang 等ACL 2023 · 被引用 8 次
- Prompts Can Play Lottery Tickets Well: Achieving Lifelong Information Extraction via Lottery Prompt TuningZujie Liang, Feng Wei, Yin Jie, Yuxi Qian 等ACL 2023 · 被引用 8 次
- Learn or Recall? Revisiting Incremental Learning with Pre-trained Language ModelsJunhao Zheng, Shengjie Qiu, Qianli MaACL 2024
它引用的顶会 Paper8
- Dice Loss for Data-imbalanced NLP TasksXiaoya Li, Xiaofei Sun, Yuxian Meng, Junjun Liang 等ACL 2020 · 被引用 575 次
- Causal Intervention for Weakly-Supervised Semantic SegmentationDong Zhang, Hanwang Zhang, Jinhui Tang, Xian-Sheng Hua 等NeurIPS 2020 · 被引用 563 次
- Long-Tailed Classification by Keeping the Good and Removing the Bad Momentum Causal EffectKaihua Tang, Jianqiang Huang, Hanwang ZhangNeurIPS 2020 · 被引用 533 次
- Uncovering Main Causalities for Long-tailed Information ExtractionGuoshun Nan, Jiaqi Zeng, Rui Qiao, Zhijiang Guo 等EMNLP 2021 · 被引用 39 次
- Counterfactual Off-Policy Training for Neural Dialogue GenerationQingfu Zhu, Wei-Nan Zhang, Ting Liu, William Yang WangEMNLP 2020 · 被引用 18 次
相关 Paper
- Unify Named Entity Recognition Scenarios via Contrastive Real-Time Updating PrototypeYanhe Liu, Peng Wang, Wenjun Ke, Guozheng Li 等AAAI 2024 · 被引用 7 次
- Few-Shot Class-Incremental Learning for Named Entity RecognitionRui Wang, Tong Yu, Handong Zhao, Sungchul Kim 等ACL 2022 · 被引用 26 次
- Continual Learning for Named Entity RecognitionNatawut Monaikul, Giuseppe Castellucci, Simone Filice, Oleg RokhlenkoAAAI 2021 · 被引用 84 次
- Prototypical Replay with Old-class Focusing Knowledge Distillation for Incremental Named Entity RecognitionZesheng Liu, Qiannan Zhu, Cuiping Li, Hong ChenAAAI 2025
- CafeBoost: Causal Feature Boost to Eliminate Task-Induced Bias for Class Incremental LearningBenliu Qiu, Hongliang Li, Haitao Wen, Heqian Qiu 等CVPR 2023
