ECENet: Explainable and Context-Enhanced Network for Muti-modal Fact verification
Fanrui Zhang, Jiawei Liu, Qiang Zhang, Esther Sun, Jingyi Xie, Zheng-Jun Zha
Abstract
Recently, falsified claims incorporating both text and images have been disseminated more effectively than those containing text alone, raising significant concerns for multi-modal fact verification. Existing research makes contributions to multi-modal feature extraction and interaction, but fails to fully utilize and enhance the valuable and intricate semantic relationships between distinct features. Moreover, most detectors merely provide a single outcome judgment and lack an inference process or explanation. Taking these factors into account, we propose a novel Explainable and Context-Enhanced Network (ECENet) for multi-modal fact verification, making the first attempt to integrate multi-clue feature extraction, multi-level feature reasoning, and justification (explanation) generation within a unified framework. Specifically, we propose an Improved Coarse- and Fine-grained Attention Network, equipped with two types of level-grained attention mechanisms, to facilitate a comprehensive understanding of contextual information. Furthermore, we propose a novel justification generation module via deep reinforcement learning that does not require additional labels. In this module, a sentence extractor agent measures the importance between the query claim and all document sentences at each time step, selecting a suitable amount of high-scoring sentences to be rewritten as the explanation of the model. Extensive experiments demonstrate the effectiveness of the proposed method.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get c17f769c-d331-4c54-95f4-7bd9f5f68b7bCited by top-tier papers7
- Learning Discriminative Noise Guidance for Image Forgery Detection and LocalizationJiaying Zhu, Dong Li, Xueyang Fu, Gang Yang et al.AAAI 2024 · 27 citations
- Fact-R1: Towards Explainable Video Misinformation Detection with Deep ReasoningFanrui Zhang, Dian Li, Qiang Zhang, Jun Chen et al.NeurIPS 2025 · 20 citations
- "Image, Tell me your story!" Predicting the original meta-context of visual misinformationJonathan Tonglet, Marie-Francine Moens, Iryna GurevychEMNLP 2024 · 6 citations
- Automated Justification Production for Claim Veracity in Fact Checking: A Survey on Architectures and ApproachesIslam Eldifrawi, Shengrui Wang, Amine TrabelsiACL 2024 · 4 citations
- A Lottery Ticket Hypothesis Approach with Sparse Fine-tuning and MAE for Image Forgery Detection and LocalizationJiaying Zhu, Dong Li, Xueyang Fu, Gege Shi et al.AAAI 2025 · 3 citations
Related papers
- Hierarchical Semantic Enhancement Network for Multimodal Fake News DetectionQiang Zhang, Jiawei Liu, Fanrui Zhang, Jingyi Xie et al.ACM MM 2023 · 10 citations
- ESCNet: Entity-enhanced and Stance Checking Network for Multi-modal Fact-CheckingFanrui Zhang, Jiawei Liu, Jingyi Xie, Qiang Zhang et al.WWW 2024 · 18 citations
- Navigating Truth in Multimodal Fact-checking via Retrieval- and Reasoning-Enhanced Large Language ModelsFanrui Zhang, Qiang Zhang, Jianwen Sun, Chuanhao Li et al.WWW 2026
- Knowledge-Enhanced Multimodal Fake News Detection: Semantic Visual and Priority FusionQin Zhang, Jiaying Liu, Qian Tao, Zhiwei Guo et al.WWW 2026
- Hierarchical Multi-modal Contextual Attention Network for Fake News DetectionShengsheng Qian, Jinguang Wang, Jun Hu, Quan Fang et al.SIGIR 2021 · 273 citations
