Listwise Preference Diffusion Optimization for User Behavior Trajectories Prediction
Hongtao Huang, Chengkai Huang, Junda Wu, Tong Yu, Julian J. McAuley, Lina Yao
Abstract
Forecasting multi-step user behavior trajectories requires reasoning over structured preferences across future actions, a challenge overlooked by traditional sequential recommendation. This problem is critical for applications such as personalized commerce and adaptive content delivery, where anticipating a user's complete action sequence enhances both satisfaction and business outcomes. We identify an essential limitation of existing paradigms: their inability to capture global, listwise dependencies among sequence items. To address this, we formulate User Behavior Trajectory Prediction (UBTP) as a new task setting that explicitly models long-term user preferences. We introduce Listwise Preference Diffusion Optimization (LPDO), a diffusion-based training framework that directly optimizes structured preferences over entire item sequences. LPDO incorporates a Plackett-Luce supervision signal and derives a tight variational lower bound aligned with listwise ranking likelihoods, enabling coherent preference generation across denoising steps and overcoming the independent-token assumption of prior diffusion methods. To rigorously evaluate multi-step prediction quality, we propose the task-specific metric Sequential Match (SeqMatch), which measures exact trajectory agreement, and adopt Perplexity (PPL), which assesses probabilistic fidelity. Extensive experiments on real-world user behavior benchmarks demonstrate that LPDO consistently outperforms state-of-the-art baselines, establishing a new benchmark for structured preference learning with diffusion models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7f74ce56-9573-40e0-97b1-d30465092b82Cited by top-tier papers3
- PruneRAG: Confidence-Guided Query Decomposition Trees for Efficient Retrieval-Augmented GenerationShuguang Jiao, Xinyu Xiao, Yunfan Wei, Shuhan Qi et al.WWW 2026 · 2 citations
- WS-GRPO: Weakly-Supervised Group-Relative Policy Optimization for Rollout-Efficient ReasoningGagan Mundada, Zihan Huang, Rohan Surana, Sheldon Yu et al.ICML 2026
- Factorized Latent Reasoning for LLM-based RecommendationTianqi Gao, Chengkai Huang, Zihan Wang, Cao Liu et al.SIGIR 2026
Builds on12
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion-LM Improves Controllable Text GenerationXiang Lisa Li, John Thickstun, Ishaan Gulrajani, Percy Liang et al.NeurIPS 2022 · 1,546 citations
- On Sampled Metrics for Item RecommendationWalid Krichene, Steffen RendleKDD 2020 · 459 citations
- Generate What You Prefer: Reshaping Sequential Recommendation via Guided DiffusionZhengyi Yang, Jiancan Wu, Zhicai Wang, Xiang Wang et al.NeurIPS 2023 · 205 citations
- Sequential Recommendation via Stochastic Self-AttentionZiwei Fan, Zhiwei Liu, Yu Wang, Alice Wang et al.WWW 2022 · 203 citations
Related papers
- Towards Better Optimization For Listwise Preference in Diffusion ModelsJiamu Bai, Xin Yu, Meilong Xu, Weitao Lu et al.ICLR 2026 · 8 citations
- Beyond Static Diffusion: Explicitly Modeling Temporal Patterns in Sequential RecommendationYao Wu, Chengyi Liu, Wenqi Fan, Rui ZhangSIGIR 2026
- Diffused Task-Agnostic Milestone PlannerMineui Hong, Minjae Kang, Songhwai OhNeurIPS 2023 · 15 citations
- Unleashing the Potential of Diffusion Models Towards Diversified Sequential RecommendationsZhuo Cai, Shoujin Wang, Victor W. Chu, Usman Naseem et al.SIGIR 2025 · 7 citations
- Planning with Diffusion Models for Target-Oriented Dialogue SystemsHanwen Du, Bo Peng, Xia NingACL 2025
