Mirror Gradient: Towards Robust Multimodal Recommender Systems via Exploring Flat Local Minima
Shanshan Zhong, Zhongzhan Huang, Daifeng Li, Wushao Wen, Jinghui Qin, Liang Lin
Abstract
Multimodal recommender systems utilize various types of information to model user preferences and item features, helping users discover items aligned with their interests. The integration of multimodal information mitigates the inherent challenges in recommender systems, e.g., the data sparsity problem and cold-start issues. However, it simultaneously magnifies certain risks from multimodal information inputs, such as information adjustment risk and inherent noise risk. These risks pose crucial challenges to the robustness of recommendation models. In this paper, we analyze multimodal recommender systems from the novel perspective of flat local minima and propose a concise yet effective gradient strategy called Mirror Gradient (MG). This strategy can implicitly enhance the model's robustness during the optimization process, mitigating instability risks arising from multimodal information inputs. We also provide strong theoretical evidence and conduct extensive empirical experiments to show the superiority of MG across various multimodal recommendation models and benchmarks. Furthermore, we find that the proposed MG can complement existing robust training methods and be easily extended to diverse advanced recommendation models, making it a promising new and fundamental paradigm for training multimodal recommender systems. The code is released at https://github.com/Qrange-group/Mirror-Gradient .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 21ec4d7d-a4a2-4c68-998b-b87db1ff9627Cited by top-tier papers7
- Generating with Fairness: A Modality-Diffused Counterfactual Framework for Incomplete Multimodal RecommendationsJin Li, Shoujin Wang, Qi Zhang, Shui Yu et al.WWW 2025 · 26 citations
- Curriculum Conditioned Diffusion for Multimodal RecommendationYimeng Yang, Haokai Ma, Lei Meng, Shuo Xu et al.AAAI 2025 · 12 citations
- DVIB: Towards Robust Multimodal Recommender Systems via Variational Information Bottleneck DistillationWenkuan Zhao, Shanshan Zhong, Yifan Liu, Wushao Wen et al.WWW 2025 · 7 citations
- Refining Contrastive Learning and Homography Relations for Multi-Modal RecommendationShouxing Ma, Yawen Zeng, Shiqing Wu, Guandong XuACM MM 2025 · 3 citations
- Joint Similarity Item Exploration and Overlapped User Guidance for Multi-Modal Cross-Domain RecommendationWeiming Liu, Chaochao Chen, Jiahe Xu, Xinting Liao et al.WWW 2025 · 3 citations
Builds on26
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language ModelsJunnan Li, Dongxu Li, Silvio Savarese, Steven C. H. HoiICML 2023 · 7,873 citations
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- Sharpness-aware Minimization for Efficiently Improving GeneralizationPierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam NeyshaburICLR 2021 · 1,861 citations
- ASAM: Adaptive Sharpness-Aware Minimization for Scale-Invariant Learning of Deep Neural NetworksJungmin Kwon, Jeongseop Kim, Hyunseo Park, In Kwon ChoiICML 2021 · 385 citations
- Graph-Refined Convolutional Network for Multimedia Recommendation with Implicit FeedbackYinwei Wei, Xiang Wang, Liqiang Nie, Xiangnan He et al.ACM MM 2020 · 374 citations
Related papers
- Enhancing Adversarial Robustness of Multi-modal Recommendation via Modality BalancingYu Shang, Chen Gao, Jiansheng Chen, Depeng Jin et al.ACM MM 2023 · 9 citations
- MDVT: Enhancing Multimodal Recommendation with Model-Agnostic Multimodal-Driven Virtual TripletsJinfeng Xu, Zheyu Chen, Jinze Li, Shuo Yang et al.KDD 2025 · 4 citations
- I3-MRec: Invariant Learning with Information Bottleneck for Incomplete Modality RecommendationHuilin Chen, Miaomiao Cai, Fan Liu, Zhiyong Cheng et al.ACM MM 2025 · 1 citation
- Harnessing Multimodal Large Language Models for Multimodal Sequential RecommendationYuyang Ye, Zhi Zheng, Yishan Shen, Tianshu Wang et al.AAAI 2025 · 68 citations
- VI-MMRec: Similarity-Aware Training Cost-free Virtual User-Item Interactions for Multimodal RecommendationJinfeng Xu, Zheyu Chen, Shuo Yang, Jinze Li et al.KDD 2026 · 3 citations
