Preference Diffusion for Recommendation
Shuo Liu, An Zhang, Guoqing Hu, Hong Qian, Tat-Seng Chua
Abstract
Recommender systems aim to predict personalized item rankings by modeling user preference distributions derived from historical behavior data. While diffusion models (DMs) have recently gained attention for their ability to model complex distributions, current DM-based recommenders typically rely on traditional objectives such as mean squared error (MSE) or standard recommendation objectives. These approaches are either suboptimal for personalized ranking tasks or fail to exploit the full generative potential of DMs. To address these limitations, we propose Prefer-Diff, an optimization objective tailored for DM-based recommenders. PreferDiff reformulates the traditional Bayesian Personalized Ranking (BPR) objective into a log-likelihood generative framework, enabling it to effectively capture user preferences by integrating multiple negative samples. To handle the intractability, we employ variational inference, minimizing the variational upper bound. Furthermore, we replace MSE with cosine error to improve alignment with recommendation tasks, and we balance generative learning and preference modeling to enhance the training stability of DMs. PreferDiff devises three appealing properties. First, it is the first personalized ranking loss designed specifically for DM-based recommenders. Second, it improves ranking performance and accelerates convergence by effectively addressing hard negatives. Third, we establish its theoretical connection to Direct Preference Optimization (DPO), demonstrating its potential to align user preferences within a generative modeling framework. Extensive experiments across six benchmarks validate PreferDiff's superior recommendation performance. Our codes are available at https://github.com/lswhim/PreferDiff .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers11
- Listwise Preference Diffusion Optimization for User Behavior Trajectories PredictionHongtao Huang, Chengkai Huang, Junda Wu, Tong Yu et al.NeurIPS 2025 · 16 citations
- On Efficiency-Effectiveness Trade-off of Diffusion-based RecommendersWenyu Mao, Jiancan Wu, Guoqing Hu, Zhengyi Yang et al.NeurIPS 2025 · 5 citations
- Language Representation Favored Zero-Shot Cross-Domain Cognitive DiagnosisShuo Liu, Zihan Zhou, Yuanhao Liu, Jing Zhang et al.KDD 2025 · 4 citations
- Fine-tuning Multimodal Large Language Models for Product BundlingXiaohao Liu, Jie Wu, Zhulin Tao, Yunshan Ma et al.KDD 2025 · 3 citations
- Exploiting Inter-Session Information with Frequency-enhanced Dual-Path Networks for Sequential RecommendationPeng He, Yanglei Gan, Tingting Dai, Run Lin et al.AAAI 2026 · 2 citations
Builds on42
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning et al.NeurIPS 2023 · 10,924 citations
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
Related papers
- On Softmax Direct Preference Optimization for RecommendationYuxin Chen, Junfei Tan, An Zhang, Zhengyi Yang et al.NeurIPS 2024 · 126 citations
- Ranking-based Preference Optimization for Diffusion Models from Implicit User FeedbackYi-Lun Wu, Bo-Kai Ruan, Chiang Tseng, Hong-Han ShuaiNeurIPS 2025 · 3 citations
- Diffusion Recommender ModelWenjie Wang, Yiyan Xu, Fuli Feng, Xinyu Lin et al.SIGIR 2023 · 281 citations
- Fading to Grow: Growing Preference Ratios via Preference Fading Discrete Diffusion for RecommendationGuoqing Hu, An Zhang, Shuchang Liu, Wenyu Mao et al.NeurIPS 2025 · 4 citations
- Smoothed Preference Optimization via ReNoise Inversion for Aligning Diffusion Models with Varied Human PreferencesYunhong Lu, Qichao Wang, Hengyuan Cao, Xiaoyin Xu et al.ICML 2025
