SGoT-R1: Social Graph of Thought Reasoning-Enhanced Multimodal Large Language Model for Harmful Meme Detection
Xiuxian Wang, Yuting Su, Wenhui Li, Xiaowen Wang, Zhuojun Li, Anan Liu
摘要
Internet memes serve as widely distributed multimodal social content that conveys complex ideas through metaphorical expressions, often containing harmful implications that make accurate harmful meme detection an important problem. Reasoning knowledge extracted from large language models plays a crucial role in recent advances in harmful meme detection. However, these methods only perform reasoning analysis on memes from a single opinion, ignoring that memes are essentially products of group consensus, where their true meaning interpretation highly depends on the collision and aggregation process of diverse user viewpoints. To address this problem, we propose a Social Graph of Thought Reasoning Enhancement (SGoTRE) framework for harmful meme detection. The SGoTRE contains three key steps: First, through multi-agent simulation technology, we obtain diverse chains of thought that represent the parsing logic of users from different backgrounds toward memes, authentically restoring the diversity characteristics of group cognition. Second, we construct a Social Graph of Thought (SGoT) that effectively integrates multi-chain reasoning processes and structurally expresses the consensus and diversity of viewpoints among users. Finally, we utilize the SGoT for cognitive distillation, internalizing multi-opinion reasoning logic into a single multimodal large model (SGoT-R1) to achieve efficient and interpretable harmful meme detection. Experimental results show that SGoT-R1 significantly improves detection performance on mainstream datasets. Particularly on the most challenging FHM dataset, SGoT-R1 achieves an 8.9% improvement over state-of-the-art models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Prompting for Multimodal Hateful Meme ClassificationRui Cao, Roy Ka-Wei Lee, Wen-Haw Chong, Jing JiangEMNLP 2022 · 被引用 68 次
- Towards Explainable Harmful Meme Detection through Multimodal Debate between Large Language ModelsHongzhan Lin, Ziyang Luo, Wei Gao, Jing Ma 等WWW 2024 · 被引用 43 次
- Multimodal Hate Speech Detection via Cross-Domain Knowledge TransferChuanpeng Yang, Fuqing Zhu, Guihua Liu, Jizhong Han 等ACM MM 2022 · 被引用 42 次
相关 Paper
- Cognitive Distillation for Information Forensics: Towards Improved Hateful Meme DetectionXiuxian Wang, Yuting Su, Wenhui Li, Ruidong Chen 等KDD 2026
- Is Having Rationales Enough? Rethinking Knowledge Enhancement for Multimodal Hateful Meme DetectionJunyu Lu, Bo Xu, Xiaokun Zhang, Haohao Zhu 等SIGIR 2025 · 被引用 3 次
- Towards Low-Resource Harmful Meme Detection with LMM AgentsJianzhao Huang, Hongzhan Lin, Ziyan Liu, Ziyang Luo 等EMNLP 2024 · 被引用 2 次
- I know what you MEME! Understanding and Detecting Harmful Memes with Multimodal Large Language ModelsYong Zhuang, Keyan Guo, Juan Wang, Yiheng Jing 等NDSS 2025
- AdamMeme: Adaptively Probe the Reasoning Capacity of Multimodal Large Language Models on HarmfulnessZixin Chen, Hongzhan Lin, Kaixin Li, Ziyang Luo 等ACL 2025
