Multimodal Taylor Series Network for Misinformation Detection
Jiahao Sun, Chen Chen, Chunyan Hou, Yike Wu, Xiaojie Yuan
摘要
With the rapid development of the Internet and the widespread use of social media, the proliferation of multimodal misinformation combining images and text poses serious risks to societal trust, individual well-being, and the integrity of AI models trained on such data. Recently, the automatic detection multimodal misinformation has become an essential area of research. However, traditional methods often rely on hierarchical neural networks that compress and fuse modalities, potentially overlooking deeper interactions between modalities and reducing model interpretability. In this paper, we present a novel Multimodal Taylor Series (MTS) network for detecting multimodal misinformation. The MTS network leverages Taylor series expansion to explicitly capture both low-order and high-order interactions between modalities, which also enhances interpretability by decomposing the model's processing into distinct terms. Additionally, the proposed MTS network avoids exponential parameter growth and maintains linear scalability, allowing the model to effectively capture complex cross-modal correlations. Extensive experiments on three benchmark datasets demonstrate that the MTS network significantly outperforms state-of-the-art models. We will release our code after the final publication of the paper. CCS Concepts • Computing methodologies → Machine learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper11
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Mind the Gap: Understanding the Modality Gap in Multi-modal Contrastive Representation LearningWeixin Liang, Yuhui Zhang, Yongchan Kwon, Serena Yeung 等NeurIPS 2022 · 被引用 834 次
- Mining Dual Emotion for Fake News DetectionXueyao Zhang, Juan Cao, Xirong Li, Qiang Sheng 等WWW 2021 · 被引用 332 次
- Cross-modal Ambiguity Learning for Multimodal Fake News DetectionYixuan Chen, Dongsheng Li, Peng Zhang, Jie Sui 等WWW 2022 · 被引用 325 次
- Reasoning with Multimodal Sarcastic Tweets via Modeling Cross-Modality Contrast and Semantic AssociationNan Xu, Zhixiong Zeng, Wenji MaoACL 2020 · 被引用 153 次
相关 Paper
- Enhancing Multimodal Misinformation Detection by Replaying the Whole Story from Image Modality PerspectiveBing Wang, Ximing Li, Yanjun Wang, Changchun Li 等AAAI 2026 · 被引用 1 次
- TRUST-VL: An Explainable News Assistant for General Multimodal Misinformation DetectionZehong Yan, Peng Qi, Wynne Hsu, Mong-Li LeeEMNLP 2025 · 被引用 1 次
- KEN: Knowledge Augmentation and Emotion Guidance Network for Multimodal Fake News DetectionPeican Zhu, Yubo Jing, Le Cheng, Keke Tang 等ACM MM 2025 · 被引用 5 次
- Hierarchical Multi-modal Contextual Attention Network for Fake News DetectionShengsheng Qian, Jinguang Wang, Jun Hu, Quan Fang 等SIGIR 2021 · 被引用 273 次
- SIDA: Social Media Image Deepfake Detection, Localization and Explanation with Large Multimodal ModelZhenglin Huang, Jinwei Hu, Xiangtai Li, Yiwei He 等CVPR 2025
