Seq2seq is All You Need for Coreference Resolution
Wenzheng Zhang, Sam Wiseman, Karl Stratos
Abstract
Existing works on coreference resolution suggest that task-specific models are necessary to achieve state-of-the-art performance. In this work, we present compelling evidence that such models are not necessary. We finetune a pretrained seq2seq transformer to map an input document to a tagged sequence encoding the coreference annotation. Despite the extreme simplicity, our model outperforms or closely matches the best coreference systems in the literature on an array of datasets. We also propose an especially simple seq2seq approach that generates only tagged spans rather than the spans interleaved with the original text. Our analysis shows that the model size, the amount of supervision, and the choice of sequence representations are key factors in performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e5417f52-b353-4619-8e36-ae943fb0dde8Cited by top-tier papers5
- Major Entity Identification: A Generalizable Alternative to Coreference ResolutionKawshik Sundar, Shubham Toshniwal, Makarand Tapaswi, Vineet GandhiEMNLP 2024 · 4 citations
- ImCoref-CeS: An Improved Lightweight Pipeline for Coreference Resolution with LLM-based Checker-Splitter RefinementKangyang Luo, Yuzhuo Bai, Shuzheng Si, Cheng Gao et al.ACL 2026 · 1 citation
- Multimodal Coreference Resolution for Chinese Social Media Dialogues: Dataset and Benchmark ApproachXingyu Li, Chen Gong, Guohong FuACL 2025
- BOOKCOREF: Coreference Resolution at Book ScaleGiuliano Martinelli, Tommaso Bonomo, Pere-Lluís Huguet Cabot, Roberto NavigliACL 2025
- Bridging Context Gaps: Leveraging Coreference Resolution for Long Contextual UnderstandingYanming Liu, Xinyue Peng, Jiannan Cao, Shi Bo et al.ICLR 2025
Builds on5
- Multitask Prompted Training Enables Zero-Shot Task GeneralizationVictor Sanh, Albert Webson, Colin Raffel, Stephen H. Bach et al.ICLR 2022 · 1,976 citations
- ZeRO: memory optimizations toward training trillion parameter modelsSamyam Rajbhandari, Jeff Rasley, Olatunji Ruwase, Yuxiong HeSC 2020 · 852 citations
- Autoregressive Entity RetrievalNicola De Cao, Gautier Izacard, Sebastian Riedel, Fabio PetroniICLR 2021 · 200 citations
- CorefQA: Coreference Resolution as Query-based Span PredictionWei Wu, Fei Wang, Arianna Yuan, Fei Wu et al.ACL 2020 · 153 citations
- Moving on from OntoNotes: Coreference Resolution Model TransferPatrick Xia, Benjamin Van DurmeEMNLP 2021 · 23 citations
Related papers
- Maverick: Efficient and Accurate Coreference Resolution Defying Recent TrendsGiuliano Martinelli, Edoardo Barba, Roberto NavigliACL 2024 · 8 citations
- Attending to Entities for Better Text UnderstandingPengxiang Cheng, Katrin ErkAAAI 2020 · 42 citations
- Coreferential Reasoning Learning for Language RepresentationDeming Ye, Yankai Lin, Jiaju Du, Zhenghao Liu et al.EMNLP 2020 · 164 citations
- Annotating Mentions Alone Enables Efficient Domain Adaptation for Coreference ResolutionNupoor Gandhi, Anjalie Field, Emma StrubellACL 2023 · 3 citations
- Probing for Referential Information in Language ModelsIonut-Teodor Sorodoc, Kristina Gulordava, Gemma BoledaACL 2020 · 31 citations
