What's Left Unsaid? Detecting and Correcting Misleading Omissions in Multimodal News Previews
Fanxiao Li, Jiaying Wu, Tingchao Fu, Dayang Li, Herun Wan, Wei Zhou, Min-Yen Kan
Abstract
Even when factually correct, social-media news previews (image-headline pairs) can induce interpretation drift: by selectively omitting crucial context, they lead readers to form judgments that diverge from what the full article supports. This covert harm is subtler than explicit misinformation, yet remains underexplored. To address this gap, we develop a multistage pipeline that simulates preview-based and context-based understanding, enabling construction of the MM-MISLEADING benchmark. Using MM-MISLEADING, we systematically evaluate open-source LVLMs and uncover pronounced blind spots in omission-based misleadingness detection. We further propose OM-GUARD, which combines (1) Interpretation-Aware Fine-Tuning for misleadingness detection and (2) Rationale-Guided Misleading Content Correction, where explicit rationales guide headline rewriting to reduce misleading impressions. Experiments show that OM-GUARD lifts an 8B model's detection accuracy to the level of a 235B LVLM while delivering markedly stronger end-to-end correction. Further analysis shows that misleadingness usually arises from local narrative shifts, such as missing background, instead of global frame changes, and identifies image-driven cases where text-only correction fails, underscoring the need for visual interventions. 1 * Corresponding Author 1 Data and code are available at https://github.com/ fanxiao15/OMGuard .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 208997fa-cf56-4767-b6c1-c8012bc2e3aaCited by top-tier papers1
Ask how each one uses itBuilds on14
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
- Visual News: Benchmark and Challenges in News Image CaptioningFuxiao Liu, Yinghan Wang, Tianlu Wang, Vicente OrdonezEMNLP 2021 · 67 citations
- Sniffer: Multimodal Large Language Model for Explainable Out-of-Context Misinformation DetectionPeng Qi, Zehong Yan, Wynne Hsu, Mong-Li LeeCVPR 2024 · 54 citations
- A Survey of Computational Framing Analysis ApproachesMohammad Ali, Naeemul HassanEMNLP 2022 · 27 citations
- Beyond the Crowd: LLM-Augmented Community Notes for Governing Health MisinformationJiaying Wu, Zihang Fu, Haonan Wang, Fanxiao Li et al.ACL 2026 · 17 citations
Related papers
- Reasoning About the Unsaid: Misinformation Detection with Omission-Aware Graph InferenceZhengjia Wang, Danding Wang, Qiang Sheng, Jiaying Wu et al.AAAI 2026 · 2 citations
- Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language ModelsJiaying Wu, Fanxiao Li, Zihang Fu, Min-Yen Kan et al.ICLR 2026 · 9 citations
- Drifting Away from Truth: GenAI-Driven News Diversity Challenges LVLM-Based Misinformation DetectionFanxiao Li, Jiaying Wu, Tingchao Fu, Yunyun Dong et al.AAAI 2026 · 4 citations
- The Coherence Trap: When MLLM-Crafted Narratives Exploit Manipulated Visual ContextsYuchen Zhang, Yaxiong Wang, Yujiao Wu, Lianwei Wu et al.CVPR 2026 · 8 citations
- Are Rationales Necessary and Sufficient? Tuning LLMs for Explainable Misinformation DetectionBing Wang, Rui Miao, Ximing Li, Chen Shen et al.KDD 2026 · 1 citation
