Adaptive Disentangled Transformer for Sequential Recommendation
Yipeng Zhang, Xin Wang, Hong Chen, Wenwu Zhu
Abstract
Sequential recommendation aims at mining time-aware user interests through modeling sequential behaviors. Transformer, as an effective architecture designed to process sequential input data, has shown its superiority in capturing sequential relations for recommendation. Nevertheless, existing Transformer architectures lack explicit regularization for layer-wise disentanglement, which fails to take advantage of disentangled representation in recommendation and leads to suboptimal performance. In this paper, we study the problem of layer-wise disentanglement for Transformer architectures and propose the Adaptive Disentangled Transformer (ADT) framework, which is able to adaptively determine the optimal degree of disentanglement of attention heads within different layers. Concretely, we propose to encourage disentanglement by requiring the independence constraint via mutual information estimation over attention heads and employing auxiliary objectives to prevent the information from collapsing into useless noise. We further propose a progressive scheduler to adaptively adjust the weights controlling the degree of disentanglement via an evolutionary process. Extensive experiments on various real-world datasets demonstrate the effectiveness of our proposed ADT framework.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 714e2b93-ba82-4416-8f2d-a33ffbc92c61Cited by top-tier papers14
- End-to-end Learnable Clustering for Intent Learning in RecommendationYue Liu, Shihao Zhu, Jun Xia, Yingwei Ma et al.NeurIPS 2024 · 56 citations
- Disentangling ID and Modality Effects for Session-based RecommendationXiaokun Zhang, Bo Xu, Zhaochun Ren, Xiaochen Wang et al.SIGIR 2024 · 32 citations
- Curriculum Co-disentangled Representation Learning across Multiple Environments for Social RecommendationXin Wang, Zirui Pan, Yuwei Zhou, Hong Chen et al.ICML 2023 · 31 citations
- FineRec: Exploring Fine-grained Sequential RecommendationXiaokun Zhang, Bo Xu, Youlin Wu, Yuan Zhong et al.SIGIR 2024 · 26 citations
- Unsupervised Graph Neural Architecture Search with Disentangled Self-SupervisionZeyang Zhang, Xin Wang, Ziwei Zhang, Guangyao Shen et al.NeurIPS 2023 · 22 citations
Builds on15
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray et al.ICML 2021 · 6,356 citations
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 725 citations
- Intent Contrastive Learning for Sequential RecommendationYongjun Chen, Zhiwei Liu, Jia Li, Julian J. McAuley et al.WWW 2022 · 429 citations
Related papers
- Personalized Behavior-Aware Transformer for Multi-Behavior Sequential RecommendationJiajie Su, Chaochao Chen, Zibin Lin, Xi Li et al.ACM MM 2023 · 47 citations
- Dual-interest Factorization-heads Attention for Sequential RecommendationGuanyu Lin, Chen Gao, Yu Zheng, Jianxin Chang et al.WWW 2023 · 17 citations
- Why Generate When You Can Transform? Unleashing Generative Attention for Dynamic RecommendationYuli Liu, Wenjun Kong, Weizhi Ma, Cheng LuoACM MM 2025
- Multi-Grained Preference Enhanced Transformer for Multi-Behavior Sequential RecommendationChuan He, Yongchao Liu, Qiang Li, Weiqiang Wang et al.KDD 2025 · 1 citation
- FuXi-γ: Efficient Sequential Recommendation with Exponential-Power Temporal Encoder and Diagonal-Sparse Positional MechanismDezhi Yi, Wei Guo, Wenyang Cui, Wenxuan He et al.KDD 2026 · 1 citation
