Are ID Embeddings Necessary? Whitening Pre-trained Text Embeddings for Effective Sequential Recommendation
Lingzi Zhang, Xin Zhou, Zhiwei Zeng, Zhiqi Shen
摘要
Recent sequential recommendation models have combined pre-trained text embeddings of items with item ID embeddings to achieve superior recommendation performance. Despite their effectiveness, the expressive power of text features in these models remains largely unexplored. While most existing models emphasize the importance of ID embeddings in recommendations, our study takes a step further by studying sequential recommendation models that only rely on text features and do not necessitate ID embeddings. Upon examining pre- trained text embeddings experimentally, we discover that they reside in an anisotropic semantic space, with an average cosine similarity of over 0.8 between items. We also demonstrate that this anisotropic nature hinders recommendation models from effectively differentiating between item representations and leads to degenerated performance. To address this issue, we propose to employ a pre-processing step known as whitening transformation, which transforms the anisotropic text feature distribution into an isotropic Gaussian distribution. Our experiments show that whitening pre-trained text embeddings in the sequential model can significantly improve recommendation performance. However, the full whitening operation might break the potential manifold of items with similar text semantics. To preserve the original semantics while benefiting from the isotropy of the whitened text features, we introduce WhitenRec+, an ensemble approach that leverages both fully whitened and relaxed whitened item representations for effective recommendations. We further discuss and analyze the benefits of our design through experiments and proofs. Experimental results on three public benchmark datasets demonstrate that WhitenRec+ outperforms state-of-the-art methods for sequential recommendation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- AlphaFuse: Learn ID Embeddings for Sequential Recommendation in Null Space of Language EmbeddingsGuoqing Hu, An Zhang, Shuo Liu, Zhibo Cai 等SIGIR 2025 · 被引用 10 次
- Language Representations Can be What Recommenders Need: Findings and PotentialsLeheng Sheng, An Zhang, Yi Zhang, Yuxin Chen 等ICLR 2025
- ProMax: Exploring the Potential of LLM-derived Profiles with Distribution Shaping for Recommender SystemsYi Zhang, Yiwen Zhang, Kai Zheng, Tong Chen 等SIGIR 2026
- SpecTran: Spectral-Aware Transformer-based Adapter for LLM-Enhanced Sequential RecommendationYu Cui, Feng Liu, Zhaoxiang Wang, Changwang Zhang 等SIGIR 2026
它引用的顶会 Paper21
- VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised LearningAdrien Bardes, Jean Ponce, Yann LeCunICLR 2022 · 被引用 1,226 次
- Contrastive Learning for Sequential RecommendationXu Xie, Fei Sun, Zhaoyang Liu, Shiwen Wu 等ICDE 2022 · 被引用 674 次
- On the Sentence Embeddings from Pre-trained Language ModelsBohan Li, Hao Zhou, Junxian He, Mingxuan Wang 等EMNLP 2020 · 被引用 538 次
- On Sampled Metrics for Item RecommendationWalid Krichene, Steffen RendleKDD 2020 · 被引用 459 次
- Sequential Recommendation with Graph Neural NetworksJianxin Chang, Chen Gao, Yu Zheng, Yiqun Hui 等SIGIR 2021 · 被引用 435 次
相关 Paper
- Dual-View Whitening on Pre-trained Text Embeddings for Sequential RecommendationLingzi Zhang, Xin Zhou, Zhiwei Zeng, Zhiqi ShenAAAI 2024 · 被引用 17 次
- LLM2Rec: Large Language Models Are Powerful Embedding Models for Sequential RecommendationYingzhi He, Xiaohao Liu, An Zhang, Yunshan Ma 等KDD 2025 · 被引用 2 次
- Towards Universal Sequence Representation Learning for Recommender SystemsYupeng Hou, Shanlei Mu, Wayne Xin Zhao, Yaliang Li 等KDD 2022 · 被引用 245 次
- IDGenRec: LLM-RecSys Alignment with Textual ID LearningJuntao Tan, Shuyuan Xu, Wenyue Hua, Yingqiang Ge 等SIGIR 2024 · 被引用 45 次
- Generative Archetype-Grounded Item Representations for Sequential RecommendationYifan Li, Jiahong Liu, Xinni Zhang, Hao Chen 等WWW 2026
