Leveraging Multi-Token Entities in Document-Level Named Entity Recognition
Anwen Hu, Zhicheng Dou, Jian-Yun Nie, Ji-Rong Wen
Abstract
Most state-of-the-art named entity recognition systems are designed to process each sentence within a document independently. These systems are easy to confuse entity types when the context information in a sentence is not sufficient enough. To utilize the context information within the whole document, most document-level work let neural networks on their own to learn the relation across sentences, which is not intuitive enough for us humans. In this paper, we divide entities to multi-token entities that contain multiple tokens and single-token entities that are composed of a single token. We propose that the context information of multi-token entities should be more reliable in document-level NER for news articles. We design a fusion attention mechanism which not only learns the semantic relevance between occurrences of the same token, but also focuses more on occurrences belonging to multi-tokens entities. To identify multi-token entities, we design an auxiliary task namely ‘Multi-token Entity Classification’ and perform this task simultaneously with document-level NER. This auxiliary task is simplified from NER and doesn't require extra annotation. Experimental results on the CoNLL-2003 dataset and OntoNotesnbm dataset show that our model outperforms state-of-the-art sentence-level and document-level NER methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Interpretable Multi-dataset Evaluation for Named Entity RecognitionJinlan Fu, Pengfei Liu, Graham NeubigEMNLP 2020 · 49 citations
- Span Graph Transformer for Document-Level Named Entity RecognitionHongli Mao, Xian-Ling Mao, Hanlin Tang, Yuming Shang et al.AAAI 2024 · 3 citations
- Text-to-Table: A New Way of Information ExtractionXueqing Wu, Jiacheng Zhang, Hang LiACL 2022
Related papers
- A Supervised Multi-Head Self-Attention Network for Nested Named Entity RecognitionYongxiu Xu, Heyan Huang, Chong Feng, Yue HuAAAI 2021 · 39 citations
- Hierarchical Contextualized Representation for Named Entity RecognitionYing Luo, Fengshun Xiao, Hai ZhaoAAAI 2020 · 138 citations
- Why Attention? Analyze BiLSTM Deficiency and Its Remedies in the Case of NERPeng-Hsuan Li, Tsu-Jui Fu, Wei-Yun MaAAAI 2020 · 74 citations
- HIT: Nested Named Entity Recognition via Head-Tail Pair and Token InteractionYu Wang, Yun Li, Hanghang Tong, Ziye ZhuEMNLP 2020 · 36 citations
- Multi-modal Graph Fusion for Named Entity Recognition with Targeted Visual GuidanceDong Zhang, Suzhong Wei, Shoushan Li, Hanqian Wu et al.AAAI 2021 · 240 citations
