Towards Effective Extraction and Evaluation of Factual Claims
Dasha Metropolitansky, Jonathan Larson
摘要
A common strategy for fact-checking long-form content generated by Large Language Models (LLMs) is extracting simple claims that can be verified independently. Since inaccurate or incomplete claims compromise fact-checking results, ensuring claim quality is critical. However, the lack of a standardized evaluation framework impedes assessment and comparison of claim extraction methods. To address this gap, we propose a framework for evaluating claim extraction in the context of fact-checking along with automated, scalable, and replicable methods for applying this framework, including novel approaches for measuring coverage and decontextualization. We also introduce Claimify, an LLM-based claim extraction method, and demonstrate that it outperforms existing methods under our evaluation framework. A key feature of Claimify is its ability to handle ambiguity and extract claims only when there is high confidence in the correct interpretation of the source text.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- VeriTrail: Closed-Domain Hallucination Detection with TraceabilityDasha Metropolitansky, Jonathan LarsonICLR 2026 · 被引用 3 次
- DeepFact: Co-Evolving Benchmarks and Agents for Deep Research FactualityYukun Huang, Leonardo F. R. Ribeiro, Momchil Hardalov, Bhuwan Dhingra 等ACL 2026 · 被引用 2 次
- Assessing the Belief Consistency of Large Language Models on the Logical Conversation ProcessTomoki Tsujimura, Matiss Rikters, Masaki Asada, Shusaku Egami 等ACL 2026
- Beyond Static Artifacts: An Evolutionary Framework for Synthetic Claim GenerationYeqing Teng, Jiasheng Si, Shuxia Lin, Linhai Zhang 等ACL 2026
它引用的顶会 Paper1
相关 Paper
- Measuring and Enhancing Human Value Alignment in Zero-Shot Document-Level Claim ExtractionYuanzhen Hao, Desheng WuWWW 2026
- VeriFact: Enhancing Long-Form Factuality Evaluation with Refined Fact Extraction and Reference FactsXin Liu, Lechen Zhang, Sheza Munir, Yiyang Gu 等EMNLP 2025 · 被引用 1 次
- DnDScore: Decontextualization and Decomposition for Factuality Verification in Long-Form Text GenerationMiriam Wanner, Benjamin Van Durme, Mark DredzeEMNLP 2025 · 被引用 17 次
- Document-level Claim Extraction and Decontextualisation for Fact-CheckingZhenyun Deng, Michael Sejr Schlichtkrull, Andreas VlachosACL 2024 · 被引用 4 次
- AFaCTA: Assisting the Annotation of Factual Claim Detection with Reliable LLM AnnotatorsJingwei Ni, Minjing Shi, Dominik Stammbach, Mrinmaya Sachan 等ACL 2024
