Enhancing Adversarial Robustness of Multi-modal Recommendation via Modality Balancing
Yu Shang, Chen Gao, Jiansheng Chen, Depeng Jin, Huimin Ma, Yong Li
摘要
Recently multi-modal recommender systems have been widely applied in real scenarios such as e-commerce businesses. Existing multi-modal recommendation methods exploit the multi-modal content of items as auxiliary information and fuse them to boost performance. Despite the superior performance achieved by multi-modal recommendation models, there's currently no understanding of their robustness to adversarial attacks. In this work, we first identify the vulnerability of existing multi-modal recommendation models. Next, we show the key reason for such vulnerability is modality imbalance, i.e., the prediction score margin between positive and negative samples in the sensitive modality will drop dramatically facing adversarial attacks and fail to be compensated by other modalities. Finally, based on this finding we propose a novel defense method to enhance the robustness of multi-modal recommendation models through modality balancing. Specifically, we first adopt an embedding distillation to obtain a pair of content-similar but prediction-different item embeddings in the sensitive modality and calculate the score margin reflecting the modality vulnerability. Then we optimize the model to utilize the score margin between positive and negative samples in other modalities to compensate for the vulnerability. The proposed method can serve as a plug-and-play module and is flexible to be applied to a wide range of multi-modal recommendation models. Extensive experiments on two real-world datasets demonstrate that our method significantly improves the robustness of multi-modal recommendation models with nearly no performance degradation on clean data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Step Vulnerability Guided Mean Fluctuation Adversarial Attack against Conditional Diffusion ModelsHongwei Yu, Jiansheng Chen, Xinlong Ding, Yudong Zhang 等AAAI 2024 · 被引用 18 次
- VRAgent-R1: Boosting Video Recommendation with MLLM-based Agents via Reinforcement LearningSiran Chen, Boyu Chen, Yuxiao Luo, Chenyun Yu 等AAAI 2026
- From Zero to Hero: Cross-modal-enhanced Adversarial Item Promotion Attack against Multimodal Recommender SystemsMengyu Yao, Ziqi Zhang, Yifeng Cai, Junlin Liu 等USENIX Security 2026
- Sign-Aware Multimodal Graph RecommendationYahong Lian, Haotian Tian, Chunyao Song, Tingjian GeAAAI 2026
- The Hidden Risk: Membership Inference Attacks on Multimodal Federated Learning via Modality ImbalanceChang Ma, Jun Li, Kang Wei, Yipeng Zhou 等ICML 2026
它引用的顶会 Paper13
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li 等SIGIR 2020 · 被引用 4,448 次
- Graph-Refined Convolutional Network for Multimedia Recommendation with Implicit FeedbackYinwei Wei, Xiang Wang, Liqiang Nie, Xiangnan He 等ACM MM 2020 · 被引用 374 次
- Mining Latent Structures for Multimedia RecommendationJinghao Zhang, Yanqiao Zhu, Qiang Liu, Shu Wu 等ACM MM 2021 · 被引用 350 次
- Bootstrap Latent Representations for Multi-modal RecommendationXin Zhou, Hongyu Zhou, Yong Liu, Zhiwei Zeng 等WWW 2023 · 被引用 326 次
- How Dataset Characteristics Affect the Robustness of Collaborative Recommendation ModelsYashar Deldjoo, Tommaso Di Noia, Eugenio Di Sciascio, Felice Antonio MerraSIGIR 2020 · 被引用 50 次
相关 Paper
- Modality-Balanced Learning for Multimedia RecommendationJinghao Zhang, Guofan Liu, Qiang Liu, Shu Wu 等ACM MM 2024 · 被引用 21 次
- Aligning Distillation For Cold-start Item RecommendationFeiran Huang, Zefan Wang, Xiao Huang, Yufeng Qian 等SIGIR 2023 · 被引用 100 次
- Online Distillation-enhanced Multi-modal Transformer for Sequential RecommendationWei Ji, Xiangyan Liu, An Zhang, Yinwei Wei 等ACM MM 2023 · 被引用 32 次
- VENOMREC: Cross-Modal Interactive Poisoning for Targeted Promotion in Multimodal LLM Recommender SystemsGuowei Guan, Yurong Hao, Jiaming Zhang, Tiantong Wu 等ICML 2026
- DVIB: Towards Robust Multimodal Recommender Systems via Variational Information Bottleneck DistillationWenkuan Zhao, Shanshan Zhong, Yifan Liu, Wushao Wen 等WWW 2025 · 被引用 7 次
