Pareto Invariant Representation Learning for Multimedia Recommendation
Shanshan Huang, Haoxuan Li, Qingsong Li, Chunyuan Zheng, Li Liu
Abstract
Multimedia recommendation involves personalized ranking tasks, where multimedia content is usually represented using a generic encoder. However, these generic representations introduce spurious correlations that fail to reveal users' true preferences. Existing works attempt to alleviate this problem by learning invariant representations, but overlook the balance between independent and identically distributed (IID) and out-of-distribution (OOD) generalization. In this paper, we propose a framework called Pareto Invariant Representation Learning (PaInvRL) to mitigate the impact of spurious correlations from an IID-OOD multi-objective optimization perspective, by learning invariant representations (intrinsic factors that attract user attention) and variant representations (other factors) simultaneously. Specifically, PaInvRL includes three iteratively executed modules: (i) heterogeneous identification module, which identifies the heterogeneous environments to reflect distributional shifts for user-item interactions; (ii) invariant mask generation module, which learns invariant masks based on the Pareto-optimal solutions that minimize the adaptive weighted Invariant Risk Minimization (IRM) and Empirical Risk (ERM) losses; (iii) convert module, which generates both variant representations and item-invariant representations for training a multi-modal recommendation model that mitigates spurious correlations and balances the generalization performance within and cross the environmental distributions. We compare the proposed PaInvRL with state-of-the-art recommendation models on three public multimedia recommendation datasets (Movielens, Tiktok, and Kwai), and the experimental results validate the effectiveness of PaInvRL for both within-and cross-environmental learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9e4c9346-3124-42a3-a40a-d02b67b8d2dbCited by top-tier papers11
- Removing Hidden Confounding in Recommendation: A Unified Multi-Task Learning ApproachHaoxuan Li, Kunhan Wu, Chunyuan Zheng, Yanghao Xiao et al.NeurIPS 2023 · 68 citations
- Debiased Collaborative Filtering with Kernel-Based Causal BalancingHaoxuan Li, Chunyuan Zheng, Yanghao Xiao, Peng Wu et al.ICLR 2024 · 29 citations
- Relaxing the Accurate Imputation Assumption in Doubly Robust Learning for Debiased Collaborative FilteringHaoxuan Li, Chunyuan Zheng, Shuyi Wang, Kunhan Wu et al.ICML 2024 · 25 citations
- MetaCoCo: A New Few-Shot Classification Benchmark with Spurious CorrelationMin Zhang, Haoxuan Li, Fei Wu, Kun KuangICLR 2024 · 18 citations
- Be Aware of the Neighborhood Effect: Modeling Selection Bias under InterferenceHaoxuan Li, Chunyuan Zheng, Sihao Ding, Peng Wu et al.ICLR 2024 · 17 citations
Builds on28
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang et al.ICML 2021 · 1,163 citations
- Graph-Refined Convolutional Network for Multimedia Recommendation with Implicit FeedbackYinwei Wei, Xiang Wang, Liqiang Nie, Xiangnan He et al.ACM MM 2020 · 374 citations
- Mining Latent Structures for Multimedia RecommendationJinghao Zhang, Yanqiao Zhu, Qiang Liu, Shu Wu et al.ACM MM 2021 · 350 citations
- Invariant Risk Minimization GamesKartik Ahuja, Karthikeyan Shanmugam, Kush R. Varshney, Amit DhurandharICML 2020 · 289 citations
Related papers
- Invariant Representation Learning for Multimedia RecommendationXiaoyu Du, Zike Wu, Fuli Feng, Xiangnan He et al.ACM MM 2022 · 45 citations
- Learning Invariant Modality Representation for Robust Multimodal Learning from a Causal Inference PerspectiveSijie Mai, Shiqin HanACL 2026
- Heterogeneous Risk MinimizationJiashuo Liu, Zheyuan Hu, Peng Cui, Bo Li et al.ICML 2021 · 170 citations
- I3-MRec: Invariant Learning with Information Bottleneck for Incomplete Modality RecommendationHuilin Chen, Miaomiao Cai, Fan Liu, Zhiyong Cheng et al.ACM MM 2025 · 1 citation
- Invariant Language ModelingMaxime Peyrard, Sarvjeet Singh Ghotra, Martin Josifoski, Vidhan Agarwal et al.EMNLP 2022 · 8 citations
