Dissecting Span Identification Tasks with Performance Prediction
Sean Papay, Roman Klinger, Sebastian Padó
摘要
Span identification (in short, span ID) tasks such as chunking, NER, or code-switching detection, ask models to identify and classify relevant spans in a text. Despite being a staple of NLP, and sharing a common structure, there is little insight on how these tasks' properties influence their difficulty, and thus little guidance on what model families work well on span ID tasks, and why. We analyze span ID tasks via performance prediction, estimating how well neural architectures do on different tasks. Our contributions are: (a) we identify key properties of span ID tasks that can inform performance prediction; (b) we carry out a large-scale experiment on English data, building a model to predict performance for unseen span ID tasks that can support architecture choices; (c), we investigate the parameters of the meta model, yielding new insights on how model and task properties interact to affect span ID performance. We find, e.g., that span frequency is especially important for LSTMs, and that CRFs help when spans are infrequent and boundaries non-distinctive.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- PeerDA: Data Augmentation via Modeling Peer Relation for Span Identification TasksWeiwen Xu, Xin Li, Yang Deng, Wai Lam 等ACL 2023 · 被引用 6 次
- Evaluating Sequence Labeling on the basis of Information TheoryEnrique Amigó, Elena Álvarez Mellado, Julio Gonzalo, Jorge Carrillo-de-AlbornozACL 2025
相关 Paper
- Multilingual Large Language Models Are Not (Yet) Code-SwitchersRuochen Zhang, Samuel Cahyawijaya, Jan Christian Blaise Cruz, Genta Indra Winata 等EMNLP 2023 · 被引用 19 次
- Boundary Enhanced Neural Span Classification for Nested Named Entity RecognitionChuanqi Tan, Wei Qiu, Mosha Chen, Rui Wang 等AAAI 2020 · 被引用 125 次
- A Unified Generative Framework for Various NER SubtasksHang Yan, Tao Gui, Junqi Dai, Qipeng Guo 等ACL 2021
- From English to Code-Switching: Transfer Learning with Strong Morphological CluesGustavo Aguilar, Thamar SolorioACL 2020 · 被引用 1 次
- Improving Pretraining Techniques for Code-Switched NLPRicheek Das, Sahasra Ranjan, Shreya Pathak, Preethi JyothiACL 2023 · 被引用 3 次
