Curriculum Conditioned Diffusion for Multimodal Recommendation
Yimeng Yang, Haokai Ma, Lei Meng, Shuo Xu, Ruobing Xie, Xiangxu Meng
Abstract
Multimodal recommendation (MMRec) aims to integrate multimodal information of items to address the inherent data sparsity issue in collaborative-based recommendation. Traditional MMRec methods typically capture the structure-level item representations from the observed user behaviors within the multimodal graph, overlooking the potential impact of negative instances for personalized preference understanding. In light of the outstanding generative ability and step-by-step inference characteristic of Diffusion Models (DMs), we propose a Curriculum Conditioned Diffusion framework for Multimodal Recommendation (CCDRec), which precisely excavates the modality-aware distribution-level correlation among multi-modalities and elegantly integrates the reverse phase of DMs into negative sampling to highlight the most suitable instances in a curricular manner. Specifically, CCDRec proposes the Diffusion-controlled Multimodal Aligning module (DMA) to align multimodal knowledge with collaborative signals by capturing the fine-grained relationships among multi-modalities in the probabilistic distribution space. Furthermore, CCDRec designs the Negative-sensitive Diffusive Inferring module (NDI) to progressively synthesize the negative sample pool with diverse hardness to support the following knowledge-aware negative sampling. To gradually ramp up the training complexity, CCDRec further introduces a Curricular Negative Sampler (CNS) to tally the curriculum learning paradigm with the reverse phase of DMA, thereby adaptively sampling the gold-standard negative instances to enhance optimization. Extensive experiments on three datasets with four diverse backbones demonstrate the effectiveness and robustness of our CCDRec. The visualization analyses also clarify the underlying mechanism of our DMA in multimodal representation alignment and CNS in curricular negative discovery. The code and the corresponding dataset will be uploaded in the Appendix.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bd28cb3d-86ed-454b-bdc7-06a3457f0ed4Cited by top-tier papers4
- Explicit Modeling of Causal Factors and Confounders for Image ClassificationWei Wu, Lei Meng, Zhuang Qi, Zixuan Li et al.AAAI 2026
- Adaptive Diffusion-based Augmentation for RecommendationNa Li, Fanghui Sun, Yan Zou, Yangfu Zhu et al.AAAI 2026
- Multimodal-enhanced Federated Recommendation: A Group-wise Fusion ApproachChunxu Zhang, Weipeng Zhang, Guodong Long, Zhiheng Xue et al.WWW 2026
- PLUM-Net: Prototype-Induced Label Structuring for Disentangled Multimodal Representation NetworkKehan Wang, Huan Zhao, Yong Wei, Xupeng Zha et al.AAAI 2026
Builds on20
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- simple diffusion: End-to-end diffusion for high resolution imagesEmiel Hoogeboom, Jonathan Heek, Tim SalimansICML 2023 · 403 citations
- Mining Latent Structures for Multimedia RecommendationJinghao Zhang, Yanqiao Zhu, Qiang Liu, Shu Wu et al.ACM MM 2021 · 350 citations
Related papers
- Generating Difficulty-aware Negative Samples via Conditional Diffusion for Multi-modal RecommendationWenze Ma, Chenyu Sun, Yanmin Zhu, Zhaobo Wang et al.SIGIR 2025 · 1 citation
- DiffMM: Multi-Modal Diffusion Model for RecommendationYangqin Jiang, Lianghao Xia, Wei Wei, Da Luo et al.ACM MM 2024 · 92 citations
- Align-for-Fusion: Harmonizing Triple Preferences via Dual-oriented Diffusion for Cross-domain Sequential RecommendationYongfu Zha, Xinxin Dong, Haokai Ma, Yonghui Yang et al.KDD 2026 · 7 citations
- Collaborative Diffusion Models for RecommendationMengru Chen, Lianghao Xia, Yong Xu, Ronghua LuoSIGIR 2025 · 3 citations
- CD-CDR: Conditional Diffusion-based Item Generation for Cross-Domain RecommendationHanyu Li, Jiayu Li, Weizhi Ma, Peijie Sun et al.SIGIR 2025 · 7 citations
