DnDScore: Decontextualization and Decomposition for Factuality Verification in Long-Form Text Generation
Miriam Wanner, Benjamin Van Durme, Mark Dredze
Abstract
The decompose-then-verify strategy for verification of Large Language Model (LLM) generations decomposes claims that are then independently verified. Decontextualization augments text (claims) to ensure it can be verified outside of the original context, enabling reliable verification. While decomposition and decontextualization have been explored independently, their interactions in a complete system have not been investigated. Their conflicting purposes can create tensions: decomposition isolates atomic facts while decontextualization inserts relevant information. Furthermore, a decontextualized subclaim presents a challenge to the verification step: what part of the augmented text should be verified as it now contains multiple atomic facts? We conduct an evaluation of different decomposition, decontextualization, and verification strategies and find that the choice of strategy matters in the resulting factuality scores. Additionally, we introduce DnDScore, a decontextualization aware verification method which validates subclaims in the context of contextual information.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8b47a67d-941b-4949-8e9c-ee8373deec8aCited by top-tier papers1
Ask how each one uses itBuilds on8
- Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP TasksYizhong Wang, Swaroop Mishra, Pegah Alipoormolabashi, Yeganeh Kordi et al.EMNLP 2022 · 238 citations
- FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text GenerationSewon Min, Kalpesh Krishna, Xinxi Lyu, Mike Lewis et al.EMNLP 2023 · 225 citations
- Long-form factuality in large language modelsJerry Wei, Chengrun Yang, Xinying Song, Yifeng Lu et al.NeurIPS 2024 · 182 citations
- Generating Literal and Implied Subquestions to Fact-check Complex ClaimsJifan Chen, Aniruddh Sriram, Eunsol Choi, Greg DurrettEMNLP 2022 · 30 citations
- MiniCheck: Efficient Fact-Checking of LLMs on Grounding DocumentsLiyan Tang, Philippe Laban, Greg DurrettEMNLP 2024 · 26 citations
Related papers
- Optimizing Decomposition for Optimal Claim VerificationYining Lu, Noah Ziems, Hy Dang, Meng JiangACL 2025 · 5 citations
- Towards Effective Extraction and Evaluation of Factual ClaimsDasha Metropolitansky, Jonathan LarsonACL 2025 · 17 citations
- DSVD: Dynamic Self-Verify Decoding for Faithful Generation in Large Language ModelsYiQiu Guo, Yuchen Yang, Zhe Chen, Pingjie Wang et al.EMNLP 2025
- VISTA: Verification In Sequential Turn-based AssessmentAshley Lewis, Andrew Perrault, Eric Fosler-Lussier, Michael WhiteACL 2026
- PatentScore: Multi-dimensional Evaluation of LLM-Generated Patent ClaimsYongmin Yoo, Qiongkai Xu, Longbing CaoEMNLP 2025 · 1 citation
