Generating Scientific Claims for Zero-Shot Scientific Fact Checking
Dustin Wright, David Wadden, Kyle Lo, Bailey Kuehl, Arman Cohan, Isabelle Augenstein, Lucy Lu Wang
Abstract
Automated scientific fact checking is difficult due to the complexity of scientific language and a lack of significant amounts of training data, as annotation requires domain expertise. To address this challenge, we propose scientific claim generation, the task of generating one or more atomic and verifiable claims from scientific sentences, and demonstrate its usefulness in zero-shot fact checking for biomedical claims. We propose CLAIMGEN-BART, a new supervised method for generating claims supported by the literature, as well as KBIN, a novel method for generating claim negations. Additionally, we adapt an existing unsupervised entity-centric method of claim generation to biomedical claims, which we call CLAIMGEN-ENTITY. Experiments on zero-shot fact checking demonstrate that both CLAIMGEN-ENTITY and CLAIMGEN-BART, coupled with KBIN, achieve up to 90% performance of fully supervised models trained on manually annotated claims and evidence. A rigorous evaluation study demonstrates significant improvement in generated claim and negation quality over existing baselines. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers18
- FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text GenerationSewon Min, Kalpesh Krishna, Xinxi Lyu, Mike Lewis et al.EMNLP 2023 · 225 citations
- Fact-Checking Complex Claims with Program-Guided ReasoningLiangming Pan, Xiaobao Wu, Xinyuan Lu, Anh Tuan Luu et al.ACL 2023 · 45 citations
- LM vs LM: Detecting Factual Errors via Cross ExaminationRoi Cohen, May Hamri, Mor Geva, Amir GlobersonEMNLP 2023 · 41 citations
- I Don't Know: Explicit Modeling of Uncertainty with an [IDK] TokenRoi Cohen, Konstantin Dobler, Eden Biran, Gerard de MeloNeurIPS 2024 · 35 citations
- Heterogeneous Graph Reasoning for Fact Checking over Texts and TablesHaisong Gong, Weizhi Xu, Shu Wu, Qiang Liu et al.AAAI 2024 · 19 citations
Builds on6
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
- Explainable Automated Fact-Checking for Public Health ClaimsNeema Kotonya, Francesca ToniEMNLP 2020 · 10 citations
- NegatER: Unsupervised Discovery of Negatives in Commonsense Knowledge BasesTara Safavi, Jing Zhu, Danai KoutraEMNLP 2021 · 9 citations
- Fact or Fiction: Verifying Scientific ClaimsDavid Wadden, Shanchuan Lin, Kyle Lo, Lucy Lu Wang et al.EMNLP 2020 · 6 citations
- COVID-Fact: Fact Extraction and Verification of Real-World Claims on COVID-19 PandemicArkadiy Saakyan, Tuhin Chakrabarty, Smaranda MuresanACL 2021
Related papers
- Lost in Translation, Found in Spans: Identifying Claims in Multilingual Social MediaShubham Mittal, Megha Sundriyal, Preslav NakovEMNLP 2023 · 4 citations
- AFaCTA: Assisting the Annotation of Factual Claim Detection with Reliable LLM AnnotatorsJingwei Ni, Minjing Shi, Dominik Stammbach, Mrinmaya Sachan et al.ACL 2024
- NewsClaims: A New Benchmark for Claim Detection from News with Attribute KnowledgeRevanth Gangi Reddy, Sai Chetan Chinthakindi, Zhenhailong Wang, Yi R. Fung et al.EMNLP 2022 · 13 citations
- Unknown Claims: Generation of Fact-Checking Training Examples from Unstructured and Structured DataJean-Flavien Bussotti, Luca Ragazzi, Giacomo Frisoni, Gianluca Moro et al.EMNLP 2024 · 3 citations
- Matter-of-Fact: A Benchmark for Verifying the Feasibility of Literature-Supported Claims in Materials SciencePeter A. Jansen, Samiah Hassan, Ruoyao WangEMNLP 2025 · 4 citations
