Hierarchical Semantic Enhancement Network for Multimodal Fake News Detection
Qiang Zhang, Jiawei Liu, Fanrui Zhang, Jingyi Xie, Zheng-Jun Zha
Abstract
The explosion of multimodal fake news content on social media has sparked widespread concern. Existing multimodal fake news detection methods have made significant contributions to the development of this field, but fail to adequately exploit the potential semantic information of images and ignore the noise embedded in news entities, which severely limits the performance of the models. In this paper, we propose a novel Hierarchical Semantic Enhancement Network (HSEN) for multimodal fake news detection by learning text-related image semantic and precise news high-order knowledge semantic information. Specifically, to complement the image semantic information, HSEN utilizes textual entities as the prompt subject vocabulary and applies reinforcement learning to discover the optimal prompt format for generating image captions specific to the corresponding textual entities, which contain multi-level cross-modal correlation information. Moreover, HSEN extracts visual and textual entities from image and text, and identifies additional visual entities from image captions to extend image semantic knowledge. Based on that, HSEN exploits an adaptive hard attention mechanism to automatically select strongly related news entities and remove irrelevant noise entities to obtain precise high-order knowledge semantic information, while generating attention mask for guiding cross-modal knowledge interaction. Extensive experiments show that our method outperforms state-of-the-art methods.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 053eba40-eee1-431f-916e-8dcf71dc118eCited by top-tier papers7
- Fact-R1: Towards Explainable Video Misinformation Detection with Deep ReasoningFanrui Zhang, Dian Li, Qiang Zhang, Jun Chen et al.NeurIPS 2025 · 20 citations
- RaCMC: Residual-Aware Compensation Network with Multi-Granularity Constraints for Fake News DetectionXinquan Yu, Ziqi Sheng, Wei Lu, Xiangyang Luo et al.AAAI 2025 · 9 citations
- Each Fake News Is Fake in Its Own Way: An Attribution Multi-Granularity Benchmark for Multimodal Fake News DetectionHao Guo, Zihan Ma, Zhi Zeng, Minnan Luo et al.AAAI 2025 · 7 citations
- Harmfully Manipulated Images Matter in Multimodal Misinformation DetectionBing Wang, Shengsheng Wang, Changchun Li, Renchu Guan et al.ACM MM 2024 · 5 citations
- Consistent and Invariant Generalization Learning for Short-video Misinformation DetectionHanghui Guo, Weijie Shi, Mengze Li, Juncheng Li et al.ACM MM 2025 · 1 citation
Related papers
- Hierarchical Multi-modal Contextual Attention Network for Fake News DetectionShengsheng Qian, Jinguang Wang, Jun Hu, Quan Fang et al.SIGIR 2021 · 273 citations
- Reinforced Adaptive Knowledge Learning for Multimodal Fake News DetectionLitian Zhang, Xiaoming Zhang, Ziyi Zhou, Feiran Huang et al.AAAI 2024 · 54 citations
- Knowledge-Enhanced Multimodal Fake News Detection: Semantic Visual and Priority FusionQin Zhang, Jiaying Liu, Qian Tao, Zhiwei Guo et al.WWW 2026
- KEN: Knowledge Augmentation and Emotion Guidance Network for Multimodal Fake News DetectionPeican Zhu, Yubo Jing, Le Cheng, Keke Tang et al.ACM MM 2025 · 5 citations
- ECENet: Explainable and Context-Enhanced Network for Muti-modal Fact verificationFanrui Zhang, Jiawei Liu, Qiang Zhang, Esther Sun et al.ACM MM 2023 · 23 citations
