LADA-Trans-NER: Adaptive Efficient Transformer for Chinese Named Entity Recognition Using Lexicon-Attention and Data-Augmentation
Jiguo Liu, Chao Liu, Nan Li, Shihao Gao, Mingqi Liu, Dali Zhu
Abstract
Recently, word enhancement has become very popular for Chinese Named Entity Recognition (NER), reducing segmentation errors and increasing the semantic and boundary information of Chinese words. However, these methods tend to ignore the semantic relationship before and after the sentence after integrating lexical information. Therefore, the regularity of word length information has not been fully explored in various word-character fusion methods. In this work, we propose a Lexicon-Attention and Data-Augmentation (LADA) method for Chinese NER. We discuss the challenges of using existing methods in incorporating word information for NER and show how our proposed methods could be leveraged to overcome those challenges. LADA is based on a Transformer Encoder that utilizes lexicon to construct a directed graph and fuses word information through updating the optimal edge of the graph. Specially, we introduce the advanced data augmentation method to obtain the optimal representation for the NER task. Experimental results show that the augmentation done using LADA can considerably boost the performance of our NER system and achieve significantly better results than previous state-of-the-art methods and variant models in the literature on four publicly available NER datasets, namely Resume, MSRA, Weibo, and OntoNotes v4. We also observe better generalization and application to a real-world setting from LADA on multi-source complex entities.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7a26683d-1f63-405a-b1b5-b33f787d7ff9Builds on5
- A Novel Cascade Binary Tagging Framework for Relational Triple ExtractionZhepei Wei, Jianlin Su, Yue Wang, Yuan Tian et al.ACL 2020 · 610 citations
- Simplify the Usage of Lexicon in Chinese NERRuotian Ma, Minlong Peng, Qi Zhang, Zhongyu Wei et al.ACL 2020 · 286 citations
- Read, Retrospect, Select: An MRC Framework to Short Text Entity LinkingYingjie Gu, Xiaoye Qu, Zhefeng Wang, Baoxing Huai et al.AAAI 2021 · 33 citations
- Dynamic Modeling Cross- and Self-Lattice Attention Network for Chinese NERShan Zhao, Minghao Hu, Zhiping Cai, Haiwen Chen et al.AAAI 2021 · 30 citations
- MECT: Multi-Metadata Embedding based Cross-Transformer for Chinese Named Entity RecognitionShuang Wu, Xiaoning Song, Zhen-Hua FengACL 2021
Related papers
- Lexicon Enhanced Chinese Sequence Labeling Using BERT AdapterWei Liu, Xiyan Fu, Yue Zhang, Wenming XiaoACL 2021
- Entity Enhanced BERT Pre-training for Chinese NERChen Jia, Yuefeng Shi, Qinrong Yang, Yue ZhangEMNLP 2020 · 59 citations
- DASA-Trans-STM: Adaptive Efficient Transformer for Short Text Matching using Data Augmentation and Semantic AwarenessJiguo Liu, Chao Liu, Meimei Li, Nan Li et al.EMNLP 2025
- Local Additivity Based Data Augmentation for Semi-supervised NERJiaao Chen, Zhenghui Wang, Ran Tian, Zichao Yang et al.EMNLP 2020 · 45 citations
- SENCR: A Span Enhanced Two-Stage Network with Counterfactual Rethinking for Chinese NERHang Zheng, Qingsong Li, Shen Chen, Yuxuan Liang et al.AAAI 2024 · 8 citations
