Decoding Matters: Addressing Amplification Bias and Homogeneity Issue in Recommendations for Large Language Models
Keqin Bao, Jizhi Zhang, Yang Zhang, Xinyue Huo, Chong Chen, Fuli Feng
Abstract
Adapting Large Language Models (LLMs) for recommendation requires careful consideration of the decoding process, given the inherent differences between generating items and natural language. Existing approaches often directly apply LLMs' original decoding methods. However, we find these methods encounter significant challenges: 1) amplification bias-where standard length normalization inflates scores for items containing tokens with generation probabilities close to 1 (termed ghost tokens), and 2) homogeneity issue-generating multiple similar or repetitive items for a user. To tackle these challenges, we introduce a new decoding approach named Debiasing-Diversifying Decoding (D 3 ). D 3 disables length normalization for ghost tokens to alleviate amplification bias, and it incorporates a text-free assistant model to encourage tokens less frequently generated by LLMs for counteracting recommendation homogeneity. Extensive experiments on real-world datasets demonstrate the method's effectiveness in enhancing accuracy and diversity. The code is available at https://github. com/SAI990323/DecodingMatters .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1031714e-e7fd-40c4-b7c8-4eac65ae6c32Cited by top-tier papers14
- Reinforced Latent Reasoning for LLM-based RecommendationYang Zhang, Wenxin Xu, Xiaoyan Zhao, Wenjie Wang et al.ICLR 2026 · 67 citations
- Reasoning over Semantic IDs Enhances Generative RecommendationYingzhi He, Yan Sun, Junfei Tan, Yuxin Chen et al.KDD 2026 · 15 citations
- ThinkRec: Thinking-based recommendation via LLMQihang Yu, Kairui Fu, Zheqi Lv, Shengyu Zhang et al.WWW 2026 · 10 citations
- Spend Search Where It Pays: Value-Guided Structured Sampling and Optimization for Generative RecommendationJie Jiang, Yangru Huang, Zeyu Wang, Changping Wang et al.KDD 2026 · 3 citations
- Rec: Towards Large Recommender Models with ReasoningRunyang You, Yongqi Li, Xinyu Lin, Xin Zhang et al.NeurIPS 2025 · 3 citations
Builds on6
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes et al.ICLR 2020 · 4,112 citations
- Efficient Memory Management for Large Language Model Serving with PagedAttentionWoosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng et al.SOSP 2023 · 1,016 citations
- Neural Text Generation With Unlikelihood TrainingSean Welleck, Ilia Kulikov, Stephen Roller, Emily Dinan et al.ICLR 2020 · 683 citations
- Recommender Systems with Generative RetrievalShashank Rajput, Nikhil Mehta, Anima Singh, Raghunandan Hulikal Keshavan et al.NeurIPS 2023 · 474 citations
- On Sampled Metrics for Item RecommendationWalid Krichene, Steffen RendleKDD 2020 · 459 citations
Related papers
- IGD: Token Decisiveness Modeling via Information Gain in LLMs for Personalized RecommendationZijie Lin, Yang Zhang, Xiaoyan Zhao, Fengbin Zhu et al.NeurIPS 2025 · 12 citations
- Learning Decomposed Contextual Token Representations from Pretrained and Collaborative Signals for Generative RecommendationYifan Liu, Yaokun Liu, Zelin Li, Zhenrui Yue et al.SIGIR 2026
- ProMax: Exploring the Potential of LLM-derived Profiles with Distribution Shaping for Recommender SystemsYi Zhang, Yiwen Zhang, Kai Zheng, Tong Chen et al.SIGIR 2026
- Beyond Utility: Evaluating LLM as RecommenderChumeng Jiang, Jiayin Wang, Weizhi Ma, Charles L. A. Clarke et al.WWW 2025 · 22 citations
- Bridging Semantic Understanding and Popularity Bias with LLMsRenqiang Luo, Dong Zhang, Yupeng Gao, Wen Shi et al.WWW 2026
