RELiC: Retrieving Evidence for Literary Claims
Katherine Thai, Yapei Chang, Kalpesh Krishna, Mohit Iyyer
摘要
Humanities scholars commonly provide evidence for claims that they make about a work of literature (e.g., a novel) in the form of quotations from the work. We collect a large-scale dataset (RELiC) of 78K literary quotations and surrounding critical analysis and use it to formulate the novel task of literary evidence retrieval, in which models are given an excerpt of literary analysis surrounding a masked quotation and asked to retrieve the quoted passage from the set of all passages in the work. Solving this retrieval task requires a deep understanding of complex literary and linguistic phenomena, which proves challenging to methods that overwhelmingly rely on lexical and semantic similarity matching. We implement a RoBERTa-based dense passage retriever for this task that outperforms existing pretrained information retrieval baselines; however, experiments and analysis by human domain experts indicate that there is substantial room for improvement over our dense retriever.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- RankGen: Improving Text Generation with Large Ranking ModelsKalpesh Krishna, Yapei Chang, John Wieting, Mohit IyyerEMNLP 2022 · 被引用 27 次
- The Essence of Contextual Understanding in Theory of Mind: A Study on Question Answering with Story CharactersChulun Zhou, Qiujing Wang, Mo Yu, Xiaoqian Yue 等ACL 2025 · 被引用 9 次
- KRISTEVA: Close Reading as a Novel Task for Benchmarking Interpretive ReasoningPeiqi Sui, Juan Diego Rodriguez, Philippe Laban, Dean Murphy 等ACL 2025 · 被引用 6 次
- Personality Understanding of Fictional Characters during Book ReadingMo Yu, Jiangnan Li, Shunyu Yao, Wenjie Pang 等ACL 2023 · 被引用 3 次
- GEM: A Native Graph-based Index for Multi-Vector RetrievalYao Tian, Zhoujin Tian, Xi Zhao, Ruiyuan Zhang 等SIGMOD 2026 · 被引用 2 次
它引用的顶会 Paper5
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- ColBERT: Efficient and Effective Passage Search via Contextualized Late Interaction over BERTOmar Khattab, Matei ZahariaSIGIR 2020 · 被引用 1,246 次
- Dense Passage Retrieval for Open-Domain Question AnsweringVladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis 等EMNLP 2020 · 被引用 142 次
- SPECTER: Document-level Representation Learning using Citation-informed TransformersArman Cohan, Sergey Feldman, Iz Beltagy, Doug Downey 等ACL 2020 · 被引用 20 次
- Fact or Fiction: Verifying Scientific ClaimsDavid Wadden, Shanchuan Lin, Kyle Lo, Lucy Lu Wang 等EMNLP 2020 · 被引用 6 次
相关 Paper
- Unsupervised Alignment-based Iterative Evidence Retrieval for Multi-hop Question AnsweringVikas Yadav, Steven Bethard, Mihai SurdeanuACL 2020 · 被引用 3 次
- LePaRD: A Large-Scale Dataset of Judicial Citations to PrecedentRobert Mahari, Dominik Stammbach, Elliott Ash, Alex PentlandACL 2024 · 被引用 1 次
- LiteraryQA: Towards Effective Evaluation of Long-document Narrative QATommaso Bonomo, Luca Gioffré, Roberto NavigliEMNLP 2025 · 被引用 1 次
- Learning to Rank Context for Named Entity Recognition Using a Synthetic DatasetArthur Amalvy, Vincent Labatut, Richard DufourEMNLP 2023 · 被引用 6 次
- Inferential Question AnsweringJamshid Mozafari, Hamed Zamani, Guido Zuccon, Adam JatowtWWW 2026
