Entity Enhanced BERT Pre-training for Chinese NER
Chen Jia, Yuefeng Shi, Qinrong Yang, Yue Zhang
摘要
Character-level BERT pre-trained in Chinese suffers a limitation of lacking lexicon information, which shows effectiveness for Chinese NER. To integrate the lexicon into pre-trained LMs for Chinese NER, we investigate a semisupervised entity enhanced BERT pre-training method. In particular, we first extract an entity lexicon from the relevant raw text using a newword discovery method. We then integrate the entity information into BERT using Char-Entity-Transformer, which augments the selfattention using a combination of character and entity representations. In addition, an entity classification task helps inject the entity information into model parameters in pre-training. The pre-trained models are used for NER finetuning. Experiments on a news dataset and two datasets annotated by ourselves for NER in long-text show that our method is highly effective and achieves the best results.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Incorporating medical knowledge in BERT for clinical relation extractionArpita Roy, Shimei PanEMNLP 2021 · 被引用 56 次
- Unsupervised Boundary-Aware Language Model Pretraining for Chinese Sequence LabelingPeijie Jiang, Dingkun Long, Yanzhao Zhang, Pengjun Xie 等EMNLP 2022 · 被引用 9 次
- Wider & Closer: Mixture of Short-channel Distillers for Zero-shot Cross-lingual Named Entity RecognitionJun-Yu Ma, Beiduo Chen, Jia-Chen Gu, Zhenhua Ling 等EMNLP 2022 · 被引用 3 次
- Lexicon Enhanced Chinese Sequence Labeling Using BERT AdapterWei Liu, Xiyan Fu, Yue Zhang, Wenming XiaoACL 2021
- Accelerating BERT Inference for Sequence Labeling via Early-ExitXiaonan Li, Yunfan Shao, Tianxiang Sun, Hang Yan 等ACL 2021
相关 Paper
- SSMI: Semantic Similarity and Mutual Information Maximization Based Enhancement for Chinese NERPengnian Qi, Biao QinAAAI 2023 · 被引用 10 次
- Simplify the Usage of Lexicon in Chinese NERRuotian Ma, Minlong Peng, Qi Zhang, Zhongyu Wei 等ACL 2020 · 被引用 286 次
- LADA-Trans-NER: Adaptive Efficient Transformer for Chinese Named Entity Recognition Using Lexicon-Attention and Data-AugmentationJiguo Liu, Chao Liu, Nan Li, Shihao Gao 等AAAI 2023 · 被引用 8 次
- Coarse-to-Fine Pre-training for Named Entity RecognitionMengge Xue, Bowen Yu, Zhenyu Zhang, Tingwen Liu 等EMNLP 2020 · 被引用 49 次
- SENCR: A Span Enhanced Two-Stage Network with Counterfactual Rethinking for Chinese NERHang Zheng, Qingsong Li, Shen Chen, Yuxuan Liang 等AAAI 2024 · 被引用 8 次
