Fatigue-Aware Bandits for Dependent Click Models
Junyu Cao, Wei Sun, Zuo-Jun Max Shen, Markus Ettl
摘要
As recommender systems send a massive amount of content to keep users engaged, users may experience fatigue which is contributed by 1) an overexposure to irrelevant content, 2) boredom from seeing too many similar recommendations. To address this problem, we consider an online learning setting where a platform learns a policy to recommend content that takes user fatigue into account. We propose an extension of the Dependent Click Model (DCM) to describe users' behavior. We stipulate that for each piece of content, its attractiveness to a user depends on its intrinsic relevance and a discount factor which measures how many similar contents have been shown. Users view the recommended content sequentially and click on the ones that they find attractive. Users may leave the platform at any time, and the probability of exiting is higher when they do not like the content. Based on user's feedback, the platform learns the relevance of the underlying content as well as the discounting effect due to content fatigue. We refer to this learning task as “fatigue-aware DCM Bandit” problem. We consider two learning scenarios depending on whether the discounting effect is known. For each scenario, we propose a learning algorithm which simultaneously explores and exploits, and characterize its regret bound.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Modeling Attrition in Recommender Systems with Departing BanditsOmer Ben-Porat, Lee Cohen, Liu Leqi, Zachary C. Lipton 等AAAI 2022 · 被引用 14 次
- Learning to Suggest Breaks: Sustainable Optimization of Long-Term User EngagementEden Saig, Nir RosenfeldICML 2023 · 被引用 9 次
- Density-based User Representation using Gaussian Process Regression for Multi-interest Personalized RetrievalHaolun Wu, Ofer Meshi, Masrour Zoghi, Fernando Diaz 等NeurIPS 2024 · 被引用 5 次
- Product Ranking for Revenue Maximization with Multiple PurchasesRenzhe Xu, Xingxuan Zhang, Bo Li, Yafeng Zhang 等NeurIPS 2022 · 被引用 4 次
相关 Paper
- Modeling User Fatigue for Sequential RecommendationNian Li, Xin Ban, Cheng Ling, Chen Gao 等SIGIR 2024 · 被引用 10 次
- Cascading Bandits: Optimizing Recommendation Frequency in Delayed Feedback EnvironmentsDairui Wang, Junyu Cao, Yan Zhang, Wei QiNeurIPS 2023 · 被引用 2 次
- Regret in Online Recommendation SystemsKaito Ariu, Narae Ryu, Se-Young Yun, Alexandre ProutièreNeurIPS 2020 · 被引用 7 次
- Cascading Reinforcement LearningYihan Du, R. Srikant, Wei ChenICLR 2024 · 被引用 2 次
- Ready for You When You Are Back: Content-Driven Session-Based Recommendation for Continuity of ExperienceBrijraj Singh, Sonal Dabral, Niranjan PedanekarAAAI 2025
