Fatigue-Aware Bandits for Dependent Click Models
Junyu Cao, Wei Sun, Zuo-Jun Max Shen, Markus Ettl
Abstract
As recommender systems send a massive amount of content to keep users engaged, users may experience fatigue which is contributed by 1) an overexposure to irrelevant content, 2) boredom from seeing too many similar recommendations. To address this problem, we consider an online learning setting where a platform learns a policy to recommend content that takes user fatigue into account. We propose an extension of the Dependent Click Model (DCM) to describe users' behavior. We stipulate that for each piece of content, its attractiveness to a user depends on its intrinsic relevance and a discount factor which measures how many similar contents have been shown. Users view the recommended content sequentially and click on the ones that they find attractive. Users may leave the platform at any time, and the probability of exiting is higher when they do not like the content. Based on user's feedback, the platform learns the relevance of the underlying content as well as the discounting effect due to content fatigue. We refer to this learning task as “fatigue-aware DCM Bandit” problem. We consider two learning scenarios depending on whether the discounting effect is known. For each scenario, we propose a learning algorithm which simultaneously explores and exploits, and characterize its regret bound.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ca7f764e-79c7-40e1-974f-6b5e779bc963Cited by top-tier papers4
- Modeling Attrition in Recommender Systems with Departing BanditsOmer Ben-Porat, Lee Cohen, Liu Leqi, Zachary C. Lipton et al.AAAI 2022 · 14 citations
- Learning to Suggest Breaks: Sustainable Optimization of Long-Term User EngagementEden Saig, Nir RosenfeldICML 2023 · 9 citations
- Density-based User Representation using Gaussian Process Regression for Multi-interest Personalized RetrievalHaolun Wu, Ofer Meshi, Masrour Zoghi, Fernando Diaz et al.NeurIPS 2024 · 5 citations
- Product Ranking for Revenue Maximization with Multiple PurchasesRenzhe Xu, Xingxuan Zhang, Bo Li, Yafeng Zhang et al.NeurIPS 2022 · 4 citations
Related papers
- Modeling User Fatigue for Sequential RecommendationNian Li, Xin Ban, Cheng Ling, Chen Gao et al.SIGIR 2024 · 10 citations
- Cascading Bandits: Optimizing Recommendation Frequency in Delayed Feedback EnvironmentsDairui Wang, Junyu Cao, Yan Zhang, Wei QiNeurIPS 2023 · 2 citations
- Regret in Online Recommendation SystemsKaito Ariu, Narae Ryu, Se-Young Yun, Alexandre ProutièreNeurIPS 2020 · 7 citations
- Cascading Reinforcement LearningYihan Du, R. Srikant, Wei ChenICLR 2024 · 2 citations
- Ready for You When You Are Back: Content-Driven Session-Based Recommendation for Continuity of ExperienceBrijraj Singh, Sonal Dabral, Niranjan PedanekarAAAI 2025
