Are Rationales Necessary and Sufficient? Tuning LLMs for Explainable Misinformation Detection
Bing Wang, Rui Miao, Ximing Li, Chen Shen, Shaotian Yan, Changchun Li, Kaiyuan Liu, Xiaosong Yuan, Jieping Ye
摘要
The rapid spread of misinformation on social media platforms has become a formidable challenge. To mitigate its proliferation, Misinformation Detection (MD) has emerged as a critical research topic. Traditional MD approaches based on small models typically perform binary classification, e.g., real and fake, through a black-box process. Recently, the rise of Large Language Models (LLMs) has enabled explainable MD, where models generate rationales that explain their decisions, thereby enhancing transparency. Existing explainable MD methods primarily focus on crafting sophisticated prompts to elicit rationales from off-the-shelf LLMs. In this work, we propose a pipeline to fine-tune a dedicated LLM specifically for explainable MD. Our pipeline begins by collecting large-scale fact-checked articles, and then uses multiple strong LLMs to produce veracity predictions and rationales. To ensure high-quality training data, we leverage a filtering strategy that selects only the correct instances for fine-tuning. While this pipeline is intuitive and prevalent, our experiments reveal that naive filtering based solely on label correctness is insufficient in practice and suffers from two critical limitations: (1) Coarse-grained labels cause insufficient rationales: Rationales filtered solely based on binary labels are insufficient to adequately support their decisions; (2) Over-verification behavior causes unnecessary rationales: Stronger LLMs tend to exhibit over-verification behavior, producing excessively verbose and unnecessary rationales. To address these issues, we introduce LonsRex, a novel data synthesis pipeline to Locate Necessary and Sufficient Rationales for Explainable MD. Specifically, we propose a metric that quantifies the contribution of each verification step to the final prediction, thereby evaluating its necessity and sufficiency. Experimental results demonstrate that LonsRex improves the accuracy of baseline LLMs by approximately 22.97% and is comparable to larger LLMs. We will publicly release our 316k raw data and the filtered version by LonsRex.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper24
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Large Language Models are Zero-Shot ReasonersTakeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo 等NeurIPS 2022 · 被引用 8,168 次
- Rumor Detection on Social Media with Bi-Directional Graph Convolutional NetworksTian Bian, Xi Xiao, Tingyang Xu, Peilin Zhao 等AAAI 2020 · 被引用 773 次
- Mining Dual Emotion for Fake News DetectionXueyao Zhang, Juan Cao, Xirong Li, Qiang Sheng 等WWW 2021 · 被引用 332 次
- Guiding Pretraining in Reinforcement Learning with Large Language ModelsYuqing Du, Olivia Watkins, Zihan Wang, Cédric Colas 等ICML 2023 · 被引用 257 次
相关 Paper
- REFLEX: Self-Refining Explainable Fact-Checking via Verdict-Anchored Style ControlChuyi Kong, Wei Gao, Jing Ma, Hongzhan Lin 等ACL 2026 · 被引用 1 次
- Following Clues, Approaching the Truth: Explainable Micro-Video Rumor Detection via Chain-of-Thought ReasoningRongpei Hong, Jian Lang, Jin Xu, Zhangtao Cheng 等WWW 2025 · 被引用 15 次
- Towards Explainable Harmful Meme Detection through Multimodal Debate between Large Language ModelsHongzhan Lin, Ziyang Luo, Wei Gao, Jing Ma 等WWW 2024 · 被引用 43 次
- The Coherence Trap: When MLLM-Crafted Narratives Exploit Manipulated Visual ContextsYuchen Zhang, Yaxiong Wang, Yujiao Wu, Lianwei Wu 等CVPR 2026 · 被引用 8 次
- Retrieval Augmented Fact Verification by Synthesizing Contrastive ArgumentsZhenrui Yue, Huimin Zeng, Lanyu Shang, Yifan Liu 等ACL 2024
