PSL: Rethinking and Improving Softmax Loss from Pairwise Perspective for Recommendation
Weiqin Yang, Jiawei Chen, Xin Xin, Sheng Zhou, Binbin Hu, Yan Feng, Chun Chen, Can Wang
Abstract
Softmax Loss (SL) is widely applied in recommender systems (RS) and has demonstrated effectiveness. This work analyzes SL from a pairwise perspective, revealing two significant limitations: 1) the relationship between SL and conventional ranking metrics like DCG is not sufficiently tight; 2) SL is highly sensitive to false negative instances. Our analysis indicates that these limitations are primarily due to the use of the exponential function. To address these issues, this work extends SL to a new family of loss functions, termed Pairwise Softmax Loss (PSL), which replaces the exponential function in SL with other appropriate activation functions. While the revision is minimal, we highlight three merits of PSL: 1) it serves as a tighter surrogate for DCG with suitable activation functions; 2) it better balances data contributions; and 3) it acts as a specific BPR loss enhanced by Distributionally Robust Optimization (DRO). We further validate the effectiveness and robustness of PSL through empirical experiments. The code is available at https://github.com/Tiny-Snow/IR-Benchmark .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c0e1d6c5-8463-4d24-af36-77188d3d9e20Cited by top-tier papers11
- CompassNav: Steering From Path Imitation to Decision Understanding In NavigationLinfeng Li, Jian Zhao, Yuan Xie, Xin Tan et al.ICLR 2026 · 14 citations
- MSL: Not All Tokens Are What You Need for Tuning LLM as a RecommenderBohao Wang, Feng Liu, Jiawei Chen, Xingyu Lou et al.SIGIR 2025 · 5 citations
- Constrained Auto-Regressive Decoding Constrains Generative RetrievalShiguang Wu, Zhaochun Ren, Xin Xin, Jiyuan Yang et al.SIGIR 2025 · 3 citations
- Rejuvenating Cross-Entropy Loss in Knowledge Distillation for Recommender SystemsZhangchi Zhu, Wei ZhangICLR 2026 · 2 citations
- Talos: Optimizing Top-K Accuracy in Recommender SystemsShengjia Zhang, Weiqin Yang, Jiawei Chen, Peng Wu et al.WWW 2026 · 1 citation
Builds on19
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- Self-supervised Graph Learning for RecommendationJiancan Wu, Xiang Wang, Fuli Feng, Xiangnan He et al.SIGIR 2021 · 1,476 citations
- Model-Agnostic Counterfactual Reasoning for Eliminating Popularity Bias in Recommender SystemTianxin Wei, Fuli Feng, Jiawei Chen, Ziwei Wu et al.KDD 2021 · 246 citations
- AutoDebias: Learning to Debias for RecommendationJiawei Chen, Hande Dong, Yang Qiu, Xiangnan He et al.SIGIR 2021 · 167 citations
- On the Equivalence of Decoupled Graph Convolution Network and Label PropagationHande Dong, Jiawei Chen, Fuli Feng, Xiangnan He et al.WWW 2021 · 122 citations
Related papers
- Advancing Loss Functions in Recommender Systems: A Comparative Study with a Rényi Divergence-Based SolutionShengjia Zhang, Jiawei Chen, Changdong Li, Sheng Zhou et al.AAAI 2025 · 6 citations
- BSL: Understanding and Improving Softmax Loss for RecommendationJunkang Wu, Jiawei Chen, Jiancan Wu, Wentao Shi et al.ICDE 2024 · 11 citations
- Learning-Efficient Yet Generalizable Collaborative Filtering for Item RecommendationYuanhao Pu, Xiaolong Chen, Xu Huang, Jin Chen et al.ICML 2024 · 8 citations
- Breaking the Top-K Barrier: Advancing Top-K Ranking Metrics Optimization in Recommender SystemsWeiqin Yang, Jiawei Chen, Shengjia Zhang, Peng Wu et al.KDD 2025
- New Insights into Metric Optimization for Ranking-based RecommendationRoger Zhe Li, Julián Urbano, Alan HanjalicSIGIR 2021 · 6 citations
