An Adversarial Imitation Click Model for Information Retrieval
Xinyi Dai, Jianghao Lin, Weinan Zhang, Shuai Li, Weiwen Liu, Ruiming Tang, Xiuqiang He, Jianye Hao, Jun Wang, Yong Yu
Abstract
Modern information retrieval systems, including web search, ads placement, and recommender systems, typically rely on learning from user feedback. Click models, which study how users interact with a ranked list of items, provide a useful understanding of user feedback for learning ranking models. Constructing "right" dependencies is the key of any successful click model. However, probabilistic graphical models (PGMs) have to rely on manually assigned dependencies, and oversimplify user behaviors. Existing neural network based methods promote PGMs by enhancing the expressive ability and allowing flexible dependencies, but still suffer from exposure bias and inferior estimation. In this paper, we propose a novel framework, Adversarial Imitation Click Model (AICM), based on imitation learning. Firstly, we explicitly learn the reward function that recovers users' intrinsic utility and underlying intentions. Secondly, we model user interactions with a ranked list as a dynamic system instead of one-step click prediction, alleviating the exposure bias problem. Finally, we minimize the JS divergence through adversarial training and learn a stable distribution of click sequences, which makes AICM generalize well across different distributions of ranked lists. A theoretical analysis has indicated that AICM reduces the exposure bias from 𝑂 (𝑇 2 ) to 𝑂 (𝑇 ). Our studies on a public web search dataset show that AICM not only outperforms state-of-the-art models in traditional click metrics but also achieves superior performance in addressing the exposure bias and recovering the underlying patterns of click sequences. CCS CONCEPTS • Information systems → Users and interactive retrieval; Query log analysis.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 00a2235a-cdbd-44cb-bb8d-20522a7d8f9cCited by top-tier papers9
- ReLLa: Retrieval-enhanced Large Language Models for Lifelong Sequential Behavior Comprehension in RecommendationJianghao Lin, Rong Shan, Chenxu Zhu, Kounianhua Du et al.WWW 2024 · 151 citations
- ClickPrompt: CTR Models are Strong Prompt Generators for Adapting Language Models to CTR PredictionJianghao Lin, Bo Chen, Hangyu Wang, Yunjia Xi et al.WWW 2024 · 58 citations
- A Graph-Enhanced Click Model for Web SearchJianghao Lin, Weiwen Liu, Xinyi Dai, Weinan Zhang et al.SIGIR 2021 · 44 citations
- MAP: A Model-agnostic Pretraining Framework for Click-through Rate PredictionJianghao Lin, Yanru Qu, Wei Guo, Xinyi Dai et al.KDD 2023 · 28 citations
- GMOCAT: A Graph-Enhanced Multi-Objective Method for Computerized Adaptive TestingHangyu Wang, Ting Long, Liang Yin, Weinan Zhang et al.KDD 2023 · 20 citations
Related papers
- Cross-Positional Attention for Debiasing ClicksHonglei Zhuang, Zhen Qin, Xuanhui Wang, Michael Bendersky et al.WWW 2021 · 43 citations
- Whole Page Unbiased Learning to RankHaitao Mao, Lixin Zou, Yujia Zheng, Jiliang Tang et al.WWW 2024 · 6 citations
- Can Clicks Be Both Labels and Features?: Unbiased Behavior Feature Collection and Uncertainty-aware Learning to RankTao Yang, Chen Luo, Hanqing Lu, Parth Gupta et al.SIGIR 2022 · 23 citations
- An Offline Metric for the Debiasedness of Click ModelsRomain Deffayet, Philipp Hager, Jean-Michel Renders, Maarten de RijkeSIGIR 2023 · 7 citations
- A Deep Recurrent Survival Model for Unbiased RankingJiarui Jin, Yuchen Fang, Weinan Zhang, Kan Ren et al.SIGIR 2020 · 16 citations
