EulerFormer: Sequential User Behavior Modeling with Complex Vector Attention
Zhen Tian, Wayne Xin Zhao, Changwang Zhang, Xin Zhao, Zhongrui Ma, Ji-Rong Wen
摘要
To capture user preference, transformer models have been widely applied to model sequential user behavior data. The core of transformer architecture lies in the self-attention mechanism, which computes the pairwise attention scores in a sequence. Due to the permutation-equivariant nature, positional encoding is used to enhance the attention between token representations. In this setting, the pairwise attention scores can be derived by both semantic difference and positional difference. However, prior studies often model the two kinds of difference measurements in different ways, which potentially limits the expressive capacity of sequence modeling.
To address this issue, this paper proposes a novel transformer variant with complex vector attention, named EulerFormer, which provides a unified theoretical framework to formulate both semantic difference and positional difference. The EulerFormer involves two key technical improvements. First, it employs a new transformation function for efficiently transforming the sequence tokens into polar-form complex vectors using Euler's formula, enabling the unified modeling of both semantic and positional information in a complex rotation form. Secondly, it develops a differential rotation mechanism, where the semantic rotation angles can be controlled by an adaptation function, enabling the adaptive integration of the semantic and positional information according to the semantic contexts. Furthermore, a phase contrastive learning task is proposed to improve the isotropy of contextual representations in EulerFormer. Our theoretical framework possesses a high degree of completeness and generality (e.g., RoPE can be instantiated as a special case). It is more robust to semantic variations and possesses more superior theoretical properties (e.g., long-term decay) in principle. Extensive experiments conducted on four public datasets demonstrate the
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Unlocking the Power of Diffusion Models in Sequential Recommendation: A Simple and Effective ApproachJialei Chen, Yuanbo Xu, Yiheng JiangKDD 2025 · 被引用 3 次
- Seeing Beyond: Extrapolative Domain Adaptive Panoramic SegmentationYuanfan Zheng, Kunyu Peng, Xu Zheng, Kailun YangCVPR 2026 · 被引用 1 次
- Token-Context Attention for NLI: An Alternative to Self-AttentionXin Zhang, Victor S. ShengAAAI 2026
- Hyena Operator for Fast Sequential RecommendationJiahao Liu, Lin Li, Zhiyuan Li, Kaixi Hu 等WWW 2026
- Why Generate When You Can Transform? Unleashing Generative Attention for Dynamic RecommendationYuli Liu, Wenjun Kong, Weizhi Ma, Cheng LuoACM MM 2025
它引用的顶会 Paper17
- Deberta: decoding-Enhanced Bert with Disentangled AttentionPengcheng He, Xiaodong Liu, Jianfeng Gao, Weizhu ChenICLR 2021 · 被引用 3,729 次
- Graph Contrastive Learning with AugmentationsYuning You, Tianlong Chen, Yongduo Sui, Ting Chen 等NeurIPS 2020 · 被引用 3,042 次
- Train Short, Test Long: Attention with Linear Biases Enables Input Length ExtrapolationOfir Press, Noah A. Smith, Mike LewisICLR 2022 · 被引用 1,168 次
- Contrastive Learning for Sequential RecommendationXu Xie, Fei Sun, Zhaoyang Liu, Shiwen Wu 等ICDE 2022 · 被引用 674 次
- Improving Graph Collaborative Filtering with Neighborhood-enriched Contrastive LearningZihan Lin, Changxin Tian, Yupeng Hou, Wayne Xin ZhaoWWW 2022 · 被引用 606 次
相关 Paper
- Dual Contrastive Transformer for Hierarchical Preference Modeling in Sequential RecommendationChengkai Huang, Shoujin Wang, Xianzhi Wang, Lina YaoSIGIR 2023 · 被引用 17 次
- Learning Attribute as Explicit Relation for Sequential RecommendationGang Liu, Fan Yang, Yang Jiao, Alireza Bagheri Garakani 等KDD 2025 · 被引用 1 次
- PermuteFormer: Efficient Relative Position Encoding for Long SequencesPeng ChenEMNLP 2021 · 被引用 16 次
- Text Is All You Need: Learning Language Representations for Sequential RecommendationJiacheng Li, Ming Wang, Jin Li, Jinmiao Fu 等KDD 2023 · 被引用 134 次
- Multifaceted User Modeling in Recommendation: A Federated Foundation Models ApproachChunxu Zhang, Guodong Long, Hongkuan Guo, Zhaojie Liu 等AAAI 2025 · 被引用 3 次
