Flip is Better than Noise: Unbiased Interest Generation for Multimedia Recommendation
Yue He, Jingxi Xie, Fengling Li, Lei Zhu, Jingjing Li
Abstract
Generative diffusion model approaches have achieved remarkable success in multimodal recommendation by generating latent user interest interaction graphs. However, current diffusion methods based on Gaussian noise introduce uncertain interest bias noise. This noise not only disrupts the original user-item interaction bipartite graph structure but also undermines the model's ability to accurately capture user interest preferences. To address these challenges, we propose Unbiased Interest Generation for Multimodal Recommendation (GenRec). Our approach aims to generate valid latent user interests while non-invasively preserving the original interest graph structure. We innovatively introduce the Multi-Modal Interest Generaction Module. During the forward process, we simulate user interest state transitions using ''forward flipping.'' In the reverse stage, we generate binary interaction graphs following a Bernoulli distribution. To further mitigate the random uncertainty during the generation process, we design a Multi-Modal Interest Debiase Module. By constructing a multimodal interest clustering space and using user interest hashing, we correct and enhance the generated interest graphs. Finally, the Multi-modal High-Order Graph Learning Optimization is employed to capture high-order interaction information between users and items. A contrastive learning loss function is used for model optimization. We conduct extensive experiments on different real-world open datasets from the industrial sector. Compared with the state-of-the-art DiffMM, our method significantly improves NDCG@20 by 7.3% on the TikTok dataset and boosts Recall@20 by 4.2% on the Sports dataset. The experimental results validate the effectiveness of GenRec. The code is publicly available at https://github.com/orangeheyue/GenRec-V1.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 66cf729b-f568-442d-8bf8-e93b9d3a5b9aRelated papers
- DiffMM: Multi-Modal Diffusion Model for RecommendationYangqin Jiang, Lianghao Xia, Wei Wei, Da Luo et al.ACM MM 2024 · 92 citations
- Diffusion Recommender ModelWenjie Wang, Yiyan Xu, Fuli Feng, Xinyu Lin et al.SIGIR 2023 · 281 citations
- Adaptive User Dynamic Interest Guidance for Generative Sequential RecommendationKai Zhu, Jing Li, Jia Wu, Yue He et al.SIGIR 2025 · 1 citation
- Generating Difficulty-aware Negative Samples via Conditional Diffusion for Multi-modal RecommendationWenze Ma, Chenyu Sun, Yanmin Zhu, Zhaobo Wang et al.SIGIR 2025 · 1 citation
- DMMD4SR: Diffusion Model-based Multi-level Multimodal Denoising for Sequential RecommendationWeihai Lu, Li YinACM MM 2025 · 11 citations
