Multimodality Invariant Learning for Multimedia-Based New Item Recommendation
Haoyue Bai, Le Wu, Min Hou, Miaomiao Cai, Zhuangzhuang He, Yuyang Zhou, Richang Hong, Meng Wang
Abstract
Multimedia-based recommendation provides personalized item suggestions by learning the content preferences of users. With the proliferation of digital devices and APPs, a huge number of new items are created rapidly over time. How to quickly provide recommendations for new items at the inference time is challenging. What's worse, real-world items exhibit varying degrees of modality missing (e.g., many short videos are uploaded without text descriptions). Though many efforts have been devoted to multimedia-based recommendations, they either could not deal with new multimedia items or assumed the modality completeness in the modeling process.
In this paper, we highlight the necessity of tackling the modality missing issue for new item recommendation. We argue that users' inherent content preference is stable and better kept invariant to arbitrary modality missing environments. Therefore, we approach this problem from a novel perspective of invariant learning. However, how to construct environments from finite user behavior training data to generalize any modality missing is challenging. To tackle this issue, we propose a novel Multimodality
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers15
- Generating with Fairness: A Modality-Diffused Counterfactual Framework for Incomplete Multimodal RecommendationsJin Li, Shoujin Wang, Qi Zhang, Shui Yu et al.WWW 2025 · 26 citations
- MixRec: Individual and Collective Mixing Empowers Data Augmentation for Recommender SystemsYi Zhang, Yiwen ZhangWWW 2025 · 13 citations
- Structured Spectral Reasoning for Frequency-Adaptive Multimodal RecommendationWei Yang, Rui Zhong, Yiqun Chen, Chi Lu et al.NeurIPS 2025 · 10 citations
- Invariance Matters: Empowering Social Recommendation via Graph Invariant LearningYonghui Yang, Le Wu, Yuxin Liao, Zhuangzhuang He et al.SIGIR 2025 · 10 citations
- Disentangling and Generating Modalities for Recommendation in Missing Modality ScenariosJiwan Kim, Hongseok Kang, Sein Kim, Kibum Kim et al.SIGIR 2025 · 8 citations
Builds on15
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang et al.ICML 2021 · 1,163 citations
- Environment Inference for Invariant LearningElliot Creager, Jörn-Henrik Jacobsen, Richard S. ZemelICML 2021 · 454 citations
- Mining Latent Structures for Multimedia RecommendationJinghao Zhang, Yanqiao Zhu, Qiang Liu, Shu Wu et al.ACM MM 2021 · 350 citations
- Contrastive Learning for Cold-Start RecommendationYinwei Wei, Xiang Wang, Qi Li, Liqiang Nie et al.ACM MM 2021 · 321 citations
- Multi-Modal Self-Supervised Learning for RecommendationWei Wei, Chao Huang, Lianghao Xia, Chuxu ZhangWWW 2023 · 256 citations
Related papers
- GoRec: A Generative Cold-start Recommendation FrameworkHaoyue Bai, Min Hou, Le Wu, Yonghui Yang et al.ACM MM 2023 · 23 citations
- Contrastive Intra- and Inter-Modality Generation for Enhancing Incomplete Multimedia RecommendationZhenghong Lin, Yanchao Tan, Yunfei Zhan, Weiming Liu et al.ACM MM 2023 · 27 citations
- Joint Similar User Exploration and Informative Behavior Guidance for Multi-Modal New Item RecommendationJianye Xie, Lianyong Qi, Weiming Liu, Anqi Wang et al.WWW 2026
- I3-MRec: Invariant Learning with Information Bottleneck for Incomplete Modality RecommendationHuilin Chen, Miaomiao Cai, Fan Liu, Zhiyong Cheng et al.ACM MM 2025 · 1 citation
- How to Learn Item Representation for Cold-Start Multimedia Recommendation?Xiaoyu Du, Xiang Wang, Xiangnan He, Zechao Li et al.ACM MM 2020 · 64 citations
