Beyond True or False: Retrieval-Augmented Hierarchical Analysis of Nuanced Claims
Priyanka Kargupta, Runchu Tian, Jiawei Han
Abstract
Claims made by individuals or entities are oftentimes nuanced and cannot be clearly labeled as entirely "true" or false"-as is frequently the case with scientific and political claims. However, a claim (e.g., "vaccine A is better than vaccine B") can be dissected into its integral aspects and sub-aspects (e.g., efficacy, safety, distribution), which are individually easier to validate. This enables a more comprehensive, structured response that provides a well-rounded perspective on a given problem while also allowing the reader to prioritize specific angles of interest within the claim (e.g., safety towards children). Thus, we propose CLAIMSPECT, a retrieval-augmented generation-based framework for automatically constructing a hierarchy of aspects typically considered when addressing a claim and enriching them with corpusspecific perspectives. This structure hierarchically partitions an input corpus to retrieve relevant segments, which assist in discovering new sub-aspects. Moreover, these segments enable the discovery of varying perspectives towards an aspect of the claim (e.g., support, neutral, or oppose) and their respective prevalence (e.g., "how many biomedical papers believe vaccine A is more transportable than B?"). We apply CLAIMSPECT to a wide variety of real-world scientific and political claims featured in our constructed dataset, showcasing its robustness and accuracy in deconstructing a nuanced claim and representing perspectives within a corpus. Through real-world case studies and human evaluation, we validate its effectiveness over multiple baselines. * Equal contribution. Claim: "Vaccine A is better than Vaccine B" Efficacy Distribution Safety Safety for Children Safety for Elderly Affirmative (80% of papers): A has a lower rate of severe allergic reactions in adults than B. Opposition (20% of papers): B has a lower rate of blood clotting incidents in adults than A.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4172d9ea-4ea1-4aaf-bd68-afc0d638fac6Builds on3
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni et al.NeurIPS 2020 · 19,162 citations
- TELEClass: Taxonomy Enrichment and LLM-Enhanced Hierarchical Text Classification with Minimal SupervisionYunyi Zhang, Ruozhen Yang, Xueqiang Xu, Rui Li et al.WWW 2025 · 53 citations
- Assessing "Implicit" Retrieval Robustness of Large Language ModelsXiaoyu Shen, Rexhina Blloshmi, Dawei Zhu, Jiahuan Pei et al.EMNLP 2024 · 3 citations
Related papers
- Document-level Claim Extraction and Decontextualisation for Fact-CheckingZhenyun Deng, Michael Sejr Schlichtkrull, Andreas VlachosACL 2024 · 4 citations
- Explainable Automated Fact-Checking for Public Health ClaimsNeema Kotonya, Francesca ToniEMNLP 2020 · 10 citations
- Finding Needles in Document Haystacks: Augmenting Serendipitous Claim Retrieval WorkflowsMoritz Dück, Steffen Holter, Robin Shing Moon Chan, Rita Sevastjanova et al.CHI 2025 · 3 citations
- Generating Literal and Implied Subquestions to Fact-check Complex ClaimsJifan Chen, Aniruddh Sriram, Eunsol Choi, Greg DurrettEMNLP 2022 · 30 citations
- AFaCTA: Assisting the Annotation of Factual Claim Detection with Reliable LLM AnnotatorsJingwei Ni, Minjing Shi, Dominik Stammbach, Mrinmaya Sachan et al.ACL 2024
