Major Entity Identification: A Generalizable Alternative to Coreference Resolution
Kawshik Sundar, Shubham Toshniwal, Makarand Tapaswi, Vineet Gandhi
Abstract
The limited generalization of coreference resolution (CR) models has been a major bottleneck in the task's broad application. Prior work has identified annotation differences, especially for mention detection, as one of the main reasons for the generalization gap and proposed using additional annotated target domain data. Rather than relying on this additional annotation, we propose an alternative referential task, Major Entity Identification (MEI), where we: (a) assume the target entities to be specified in the input, and (b) limit the task to only the frequent entities. Through extensive experiments, we demonstrate that MEI models generalize well across domains on multiple datasets with supervised models and LLM-based few-shot prompting. Additionally, MEI fits the classification framework, which enables the use of robust and intuitive classification-based metrics. Finally, MEI is also of practical use as it allows a user to search for all mentions of a particular entity or a group of entities of interest. 1 LitBank FantasyCoref Statistics CR MEI CR MEI
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext da8132c1-bfe7-41e5-b523-884ecebcd84dCited by top-tier papers2
- ImCoref-CeS: An Improved Lightweight Pipeline for Coreference Resolution with LLM-based Checker-Splitter RefinementKangyang Luo, Yuzhuo Bai, Shuzheng Si, Cheng Gao et al.ACL 2026 · 1 citation
- BOOKCOREF: Coreference Resolution at Book ScaleGiuliano Martinelli, Tommaso Bonomo, Pere-Lluís Huguet Cabot, Roberto NavigliACL 2025
Builds on7
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Large language models are few-shot clinical information extractorsMonica Agrawal, Stefan Hegselmann, Hunter Lang, Yoon Kim et al.EMNLP 2022 · 285 citations
- Moving on from OntoNotes: Coreference Resolution Model TransferPatrick Xia, Benjamin Van DurmeEMNLP 2021 · 23 citations
- Adapting Coreference Resolution Models through Active LearningMichelle Yuan, Patrick Xia, Chandler May, Benjamin Van Durme et al.ACL 2022 · 20 citations
- Conundrums in Entity Coreference Resolution: Making Sense of the State of the ArtJing Lu, Vincent NgEMNLP 2020 · 13 citations
Related papers
- Annotating Mentions Alone Enables Efficient Domain Adaptation for Coreference ResolutionNupoor Gandhi, Anjalie Field, Emma StrubellACL 2023 · 3 citations
- Few-Shot Fine-Grained Entity Typing with Automatic Label Interpretation and Instance GenerationJiaxin Huang, Yu Meng, Jiawei HanKDD 2022 · 17 citations
- OneNet: A Fine-Tuning Free Framework for Few-Shot Entity Linking via Large Language Model PromptingXukai Liu, Ye Liu, Kai Zhang, Kehang Wang et al.EMNLP 2024 · 7 citations
- PMRC: Prompt-Based Machine Reading Comprehension for Few-Shot Named Entity RecognitionJin Huang, Danfeng Yan, Yuanqiang CaiAAAI 2024 · 2 citations
- Effective Few-Shot Named Entity Linking by Meta-LearningXiuxing Li, Zhenyu Li, Zhengyan Zhang, Ning Liu et al.ICDE 2022 · 14 citations
