Document-level Claim Extraction and Decontextualisation for Fact-Checking
Zhenyun Deng, Michael Sejr Schlichtkrull, Andreas Vlachos
摘要
Selecting which claims to check is a timeconsuming task for human fact-checkers, especially from documents consisting of multiple sentences and containing multiple claims. However, existing claim extraction approaches focus more on identifying and extracting claims from individual sentences, e.g., identifying whether a sentence contains a claim or the exact boundaries of the claim within a sentence. In this paper, we propose a method for documentlevel claim extraction for fact-checking, which aims to extract check-worthy claims from documents and decontextualise them so that they can be understood out of context. Specifically, we first recast claim extraction as extractive summarization in order to identify central sentences from documents, then rewrite them to include necessary context from the originating document through sentence decontextualisation. Evaluation with both automatic metrics and a fact-checking professional shows that our method is able to extract check-worthy claims from documents more accurately than previous work, while also improving evidence retrieval.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- When Misinformation Speaks and Converses: Rethinking Fact-Checking in Audio PlatformsChaewan Chun, Delvin Ce Zhang, Dongwon LeeACL 2026
- Improving Zero-shot Sentence Decontextualisation with Content Selection and PlanningZhenyun Deng, Yulong Chen, Andreas VlachosEMNLP 2025
- TracSum: A New Benchmark for Aspect-Based Summarization with Sentence-Level Traceability in Medical DomainBohao Chu, Meijie Li, Sameh Frihat, Chengyu Gu 等EMNLP 2025
- The Psychology of Falsehood: A Human-Centric Survey of Misinformation DetectionArghodeep Nandi, Megha Sundriyal, Euna Mehnaz Khan, Jikai Sun 等EMNLP 2025
它引用的顶会 Paper3
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Empowering the Fact-checkers! Automatic Identification of Claim Spans on TwitterMegha Sundriyal, Atharva Kulkarni, Vaibhav Pulastya, Md. Shad Akhtar 等EMNLP 2022 · 被引用 12 次
相关 Paper
- Measuring and Enhancing Human Value Alignment in Zero-Shot Document-Level Claim ExtractionYuanzhen Hao, Desheng WuWWW 2026
- Towards Effective Extraction and Evaluation of Factual ClaimsDasha Metropolitansky, Jonathan LarsonACL 2025 · 被引用 17 次
- Is the Top Still Spinning? Evaluating Subjectivity in Narrative UnderstandingMelanie Subbiah, Akankshya Mishra, Grace Kim, Liyan Tang 等EMNLP 2025
- Varifocal Question Generation for Fact-checkingNedjma Ousidhoum, Zhangdie Yuan, Andreas VlachosEMNLP 2022 · 被引用 11 次
- DnDScore: Decontextualization and Decomposition for Factuality Verification in Long-Form Text GenerationMiriam Wanner, Benjamin Van Durme, Mark DredzeEMNLP 2025 · 被引用 17 次
