Cognitive Distillation for Information Forensics: Towards Improved Hateful Meme Detection
Xiuxian Wang, Yuting Su, Wenhui Li, Ruidong Chen, Zhuojun Li, Anan Liu
摘要
Internet meme is a mainstream multimodal content form that often conveys users' viewpoints metaphorically and may also carry implicit hateful information, making hateful meme detection an extremely challenging task. Recent evidence-enhanced methods achieve remarkable research progress by retrieving meme-related evidence samples for analogical reasoning. However, most existing methods rely solely on image-text semantic similarity for information forensics, ignoring the key characteristic of meme that their meanings are shaped by the collective cognition of diverse users. To address this issue, we propose a Cognition-driven Information Forensics (CIF) framework for hateful meme detection, which integrates the collective cognition of diverse users into the entire process of information forensics and hateful content detection. This framework consists of three core modules: (1) Diverse User Comments Acquisition: it adopts a multi-agent comment simulation system to generate diverse user comments for capturing the characteristics of collective cognition; (2) Feature to Cognition Alignment Distillation: it maps visual and textual features to a semantic space aligned with collective cognition; (3) Dual-Level Association Learning: it strengthens the mining of implicit hateful clues by modeling the intra-sample and inter-sample correlation relationships. Experimental results demonstrate that the CIF framework significantly improves detection performance on mainstream datasets. In particular, on the most challenging FHM dataset, its detection performance is 4.2% higher than that of the current state-of-the-art model. Caution: Contains academic discussions of hatespeech; viewer discretion advised.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- SGoT-R1: Social Graph of Thought Reasoning-Enhanced Multimodal Large Language Model for Harmful Meme DetectionXiuxian Wang, Yuting Su, Wenhui Li, Xiaowen Wang 等AAAI 2026
- Is Having Rationales Enough? Rethinking Knowledge Enhancement for Multimodal Hateful Meme DetectionJunyu Lu, Bo Xu, Xiaokun Zhang, Haohao Zhu 等SIGIR 2025 · 被引用 3 次
- Disentangling Hate in Online MemesRoy Ka-Wei Lee, Rui Cao, Ziqing Fan, Jing Jiang 等ACM MM 2021 · 被引用 85 次
- Improving Hateful Meme Detection through Retrieval-Guided Contrastive LearningJingbiao Mei, Jinghong Chen, Weizhe Lin, Bill Byrne 等ACL 2024 · 被引用 13 次
- Uncertainty-Guided Modal Rebalance for Hateful Memes DetectionChuanpeng Yang, Yaxin Liu, Fuqing Zhu, Jizhong Han 等ACL 2024
