Temporally-Informed Analysis of Named Entity Recognition
Shruti Rijhwani, Daniel Preotiuc-Pietro
摘要
Natural language processing models often have to make predictions on text data that evolves over time as a result of changes in language use or the information described in the text. However, evaluation results on existing data sets are seldom reported by taking the timestamp of the document into account. We analyze and propose methods that make better use of temporally-diverse training data, with a focus on the task of named entity recognition. To support these experiments, we introduce a novel data set of English tweets annotated with named entities. 1 We empirically demonstrate the effect of temporal drift on performance, and how the temporal information of documents can be used to obtain better models compared to those that disregard temporal information. Our analysis gives insights into why this information is useful, in the hope of informing potential avenues of improvement for named entity recognition as well as other NLP tasks under similar experimental setups.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Mind the Gap: Assessing Temporal Generalization in Neural Language ModelsAngeliki Lazaridou, Adhiguna Kuncoro, Elena Gribovskaya, Devang Agrawal 等NeurIPS 2021 · 被引用 315 次
- Data Augmentation for Cross-Domain Named Entity RecognitionShuguang Chen, Gustavo Aguilar, Leonardo Neves, Thamar SolorioEMNLP 2021 · 被引用 39 次
- Multi-Domain Named Entity Recognition with Genre-Aware and Agnostic InferenceJing Wang, Mayank Kulkarni, Daniel Preotiuc-PietroACL 2020 · 被引用 31 次
- Event Occurrence Date Estimation based on Multivariate Time Series Analysis over Temporal Document CollectionsJiexin Wang, Adam Jatowt, Masatoshi YoshikawaSIGIR 2021 · 被引用 11 次
- Improving Temporal Generalization of Pre-trained Language Models with Lexical Semantic ChangeZhaochen Su, Zecheng Tang, Xinyan Guan, Lijun Wu 等EMNLP 2022 · 被引用 11 次
它引用的顶会 Paper1
相关 Paper
- Do CoNLL-2003 Named Entity Taggers Still Work Well in 2023?Shuheng Liu, Alan RitterACL 2023 · 被引用 9 次
- Early Discovery of Disappearing Entities in MicroblogsSatoshi Akasaki, Naoki Yoshinaga, Masashi ToyodaACL 2023 · 被引用 1 次
- Characterizing and Measuring Linguistic Dataset DriftTyler A. Chang, Kishaloy Halder, Neha Anna John, Yogarshi Vyas 等ACL 2023 · 被引用 2 次
- MEANT: Multimodal Encoder for Antecedent InformationBenjamin Irving, Annika Marie SchoeneEMNLP 2024
- Back to the Future - Temporal Adaptation of Text RepresentationsJohannes Bjerva, Wouter M. Kouw, Isabelle AugensteinAAAI 2020 · 被引用 10 次
