Sequential Recommendation via Stochastic Self-Attention
Ziwei Fan, Zhiwei Liu, Yu Wang, Alice Wang, Zahra Nazari, Lei Zheng, Hao Peng, Philip S. Yu
Abstract
Sequential recommendation models the dynamics of a user's previous behaviors in order to forecast the next item, and has drawn a lot of attention. Transformer-based approaches, which embed items as vectors and use dot-product self-attention to measure the relationship between items, demonstrate superior capabilities among existing sequential methods. However, users' real-world sequential behaviors are uncertain rather than deterministic, posing a significant challenge to present techniques. We further suggest that dot-product-based approaches cannot fully capture collaborative transitivity, which can be derived in item-item transitions inside sequences and is beneficial for cold start items. We further argue that BPR loss has no constraint on positive and sampled negative items, which misleads the optimization. We propose a novel STOchastic Self-Attention (STOSA) to overcome these issues. STOSA, in particular, embeds each item as a stochastic Gaussian distribution, the covariance of which encodes the uncertainty. We devise a novel Wasserstein Self-Attention module to characterize item-item position-wise relationships in sequences, which effectively incorporates uncertainty into model training. Wasserstein attentions also enlighten the collaborative transitivity learning as it satisfies triangle inequality. Moreover, we introduce a novel regularization term to the ranking loss, which assures the dissimilarity between positive and the negative items. Extensive experiments on five real-world benchmark datasets demonstrate the superiority of the proposed model over state-of-the-art baselines, especially on cold start items. The code is available in https://github.com/zfan20/STOSA . CCS CONCEPTS • Information systems → Recommender systems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 009757db-1e78-4a59-a283-04d5c1689ec7Cited by top-tier papers38
- Frequency Enhanced Hybrid Attention Network for Sequential RecommendationXinyu Du, Huanhuan Yuan, Pengpeng Zhao, Jianfeng Qu et al.SIGIR 2023 · 142 citations
- Sparse Meets Dense: Unified Generative Recommendations with Cascaded Sparse-Dense RepresentationsYuhao Yang, Zhi Ji, Zhaopeng Li, Yi Li et al.NeurIPS 2025 · 90 citations
- Graph Masked Autoencoder for Sequential RecommendationYaowen Ye, Lianghao Xia, Chao HuangSIGIR 2023 · 63 citations
- Meta-optimized Contrastive Learning for Sequential RecommendationXiuyuan Qin, Huanhuan Yuan, Pengpeng Zhao, Junhua Fang et al.SIGIR 2023 · 57 citations
- End-to-end Learnable Clustering for Intent Learning in RecommendationYue Liu, Shihao Zhu, Jun Xia, Yingwei Ma et al.NeurIPS 2024 · 56 citations
Builds on12
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- On Sampled Metrics for Item RecommendationWalid Krichene, Steffen RendleKDD 2020 · 459 citations
- Attentional Graph Convolutional Networks for Knowledge Concept Recommendation in MOOCs in a Heterogeneous ViewJibing Gong, Shen Wang, Jinlong Wang, Wenzheng Feng et al.SIGIR 2020 · 180 citations
- Symmetric Metric Learning with Adaptive Margin for RecommendationMingming Li, Shuai Zhang, Fuqing Zhu, Wanhui Qian et al.AAAI 2020 · 67 citations
- Probabilistic Metric Learning with Adaptive Margin for Top-K RecommendationChen Ma, Liheng Ma, Yingxue Zhang, Ruiming Tang et al.KDD 2020 · 54 citations
Related papers
- Sequential Recommendation with Relation-Aware Kernelized Self-AttentionMingi Ji, Weonyoung Joo, Kyungwoo Song, Yoon-Yeong Kim et al.AAAI 2020 · 31 citations
- Variational Self-attention Network for Sequential RecommendationJing Zhao, Pengpeng Zhao, Lei Zhao, Yanchi Liu et al.ICDE 2021 · 52 citations
- Why Generate When You Can Transform? Unleashing Generative Attention for Dynamic RecommendationYuli Liu, Wenjun Kong, Weizhi Ma, Cheng LuoACM MM 2025
- Category-aware Collaborative Sequential RecommendationRenqin Cai, Jibang Wu, Aidan San, Chong Wang et al.SIGIR 2021 · 80 citations
- Probabilistic Attention for Sequential RecommendationYuli Liu, Christian Walder, Lexing Xie, Yiqun LiuKDD 2024 · 6 citations
