Dual Cache for Long Document Neural Coreference Resolution
Qipeng Guo, Xiangkun Hu, Yue Zhang, Xipeng Qiu, Zheng Zhang
Abstract
Recent works show the effectiveness of cachebased neural coreference resolution models on long documents. These models incrementally process a long document from left to right and extract relations between mentions and entities in a cache, resulting in much lower memory and computation cost compared to computing all mentions in parallel. However, they do not handle cache misses when high-quality entities are purged from the cache, which causes wrong assignments and leads to prediction errors. We propose a new hybrid cache that integrates two eviction policies to capture global and local entities separately, and effectively reduces the aggregated cache misses up to half as before, while improving F1 score of coreference by 0.7 ∼ 5.7pt. As such, the hybrid policy can accelerate existing cache-based models and offer a new long document coreference resolution solution. Results show that our method outperforms existing methods on four benchmarks while saving up to 83% of inference time against non-cache-based models. Further, we achieve a new state-of-the-art on a long document coreference benchmark, LitBank.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- ImCoref-CeS: An Improved Lightweight Pipeline for Coreference Resolution with LLM-based Checker-Splitter RefinementKangyang Luo, Yuzhuo Bai, Shuzheng Si, Cheng Gao et al.ACL 2026 · 1 citation
- BOOKCOREF: Coreference Resolution at Book ScaleGiuliano Martinelli, Tommaso Bonomo, Pere-Lluís Huguet Cabot, Roberto NavigliACL 2025
- Mahānāma: A Unique Testbed for Literary Entity Discovery and LinkingSujoy Sarkar, Gourav Sarkar, Manoj Balaji Jagadeeshan, Jivnesh Sandhan et al.EMNLP 2025
- xCoRe: Cross-context Coreference ResolutionGiuliano Martinelli, Bruno Gatti, Roberto NavigliEMNLP 2025
Builds on3
- WinoGrande: An Adversarial Winograd Schema Challenge at ScaleKeisuke Sakaguchi, Ronan Le Bras, Chandra Bhagavatula, Yejin ChoiAAAI 2020 · 3,037 citations
- Discourse-Aware Neural Extractive Text SummarizationJiacheng Xu, Zhe Gan, Yu Cheng, Jingjing LiuACL 2020 · 264 citations
- CorefQA: Coreference Resolution as Query-based Span PredictionWei Wu, Fei Wang, Arianna Yuan, Fei Wu et al.ACL 2020 · 153 citations
Related papers
- Interpretable Coreference Resolution Evaluation Using Explicit SemanticsBruno Gatti, Giuliano Martinelli, Roberto NavigliACL 2026
- Judge Q: Trainable Queries for Optimized Information Retention in KV Cache EvictionYijun Liu, Yixuan Wang, Yuzhuang Xu, Shiyu Ji et al.AAAI 2026 · 1 citation
- NACL: A General and Effective KV Cache Eviction Framework for LLM at Inference TimeYilong Chen, Guoxia Wang, Junyuan Shang, Shiyao Cui et al.ACL 2024 · 8 citations
- Sentence-Incremental Neural Coreference ResolutionMatt Grenander, Shay B. Cohen, Mark SteedmanEMNLP 2022 · 4 citations
- RefreshKV: Updating Small KV Cache During Long-form GenerationFangyuan Xu, Tanya Goyal, Eunsol ChoiACL 2025 · 6 citations
