Why Generate When You Can Transform? Unleashing Generative Attention for Dynamic Recommendation
Yuli Liu, Wenjun Kong, Weizhi Ma, Cheng Luo
Abstract
Sequential Recommendation (SR) focuses on personalizing user experiences by predicting future preferences based on historical interactions. Transformer models, with their attention mechanisms, have become the dominant architecture in SR tasks due to their ability to capture dependencies in user behavior sequences. However, traditional attention mechanisms, where attention weights are computed through query-key transformations, are inherently linear and deterministic. This fixed approach limits their ability to account for the dynamic and non-linear nature of user preferences, leading to challenges in capturing evolving interests and subtle behavioral patterns. Given that generative models excel at capturing non-linearity and probabilistic variability, we argue that generating attention distributions offers a more flexible and expressive alternative compared to traditional attention mechanisms. To support this claim, we present a theoretical proof demonstrating that generative attention mechanisms offer greater expressiveness and stochasticity than traditional deterministic approaches. Building upon this theoretical foundation, we introduce two generative attention models for SR, each grounded in the principles of Variational Autoencoders (VAE) and Diffusion Models (DMs), respectively. These models are designed specifically to generate adaptive attention distributions that better align with variable user preferences. Extensive experiments on real-world datasets show our models significantly outperform state-of-the-art in both accuracy and diversity.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b9e86949-5b54-4b23-8985-fa26720b9f81Builds on22
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Contrastive Learning for Sequential RecommendationXu Xie, Fei Sun, Zhaoyang Liu, Shiwen Wu et al.ICDE 2022 · 674 citations
- Better Diffusion Models Further Improve Adversarial TrainingZekai Wang, Tianyu Pang, Chao Du, Min Lin et al.ICML 2023 · 300 citations
- Sequential Recommendation via Stochastic Self-AttentionZiwei Fan, Zhiwei Liu, Yu Wang, Alice Wang et al.WWW 2022 · 203 citations
- Probabilistic Transformer For Time Series AnalysisBinh Tang, David S. MattesonNeurIPS 2021 · 150 citations
Related papers
- Variational Self-attention Network for Sequential RecommendationJing Zhao, Pengpeng Zhao, Lei Zhao, Yanchi Liu et al.ICDE 2021 · 52 citations
- Probabilistic Attention for Sequential RecommendationYuli Liu, Christian Walder, Lexing Xie, Yiqun LiuKDD 2024 · 6 citations
- Adaptive User Dynamic Interest Guidance for Generative Sequential RecommendationKai Zhu, Jing Li, Jia Wu, Yue He et al.SIGIR 2025 · 1 citation
- BlossomRec: Block-level Fused Sparse Attention Mechanism for Sequential RecommendationsMengyang Ma, Xiaopeng Li, Wanyu Wang, Zhaocheng Du et al.WWW 2026 · 1 citation
- Learning Self-Modulating Attention in Continuous Time Space with Applications to Sequential RecommendationChao Chen, Haoyu Geng, Nianzu Yang, Junchi Yan et al.ICML 2021 · 12 citations
