An Adversarial Imitation Click Model for Information Retrieval
Xinyi Dai, Jianghao Lin, Weinan Zhang, Shuai Li, Weiwen Liu, Ruiming Tang, Xiuqiang He, Jianye Hao, Jun Wang, Yong Yu
摘要
Modern information retrieval systems, including web search, ads placement, and recommender systems, typically rely on learning from user feedback. Click models, which study how users interact with a ranked list of items, provide a useful understanding of user feedback for learning ranking models. Constructing "right" dependencies is the key of any successful click model. However, probabilistic graphical models (PGMs) have to rely on manually assigned dependencies, and oversimplify user behaviors. Existing neural network based methods promote PGMs by enhancing the expressive ability and allowing flexible dependencies, but still suffer from exposure bias and inferior estimation. In this paper, we propose a novel framework, Adversarial Imitation Click Model (AICM), based on imitation learning. Firstly, we explicitly learn the reward function that recovers users' intrinsic utility and underlying intentions. Secondly, we model user interactions with a ranked list as a dynamic system instead of one-step click prediction, alleviating the exposure bias problem. Finally, we minimize the JS divergence through adversarial training and learn a stable distribution of click sequences, which makes AICM generalize well across different distributions of ranked lists. A theoretical analysis has indicated that AICM reduces the exposure bias from 𝑂 (𝑇 2 ) to 𝑂 (𝑇 ). Our studies on a public web search dataset show that AICM not only outperforms state-of-the-art models in traditional click metrics but also achieves superior performance in addressing the exposure bias and recovering the underlying patterns of click sequences. CCS CONCEPTS • Information systems → Users and interactive retrieval; Query log analysis.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- ReLLa: Retrieval-enhanced Large Language Models for Lifelong Sequential Behavior Comprehension in RecommendationJianghao Lin, Rong Shan, Chenxu Zhu, Kounianhua Du 等WWW 2024 · 被引用 151 次
- ClickPrompt: CTR Models are Strong Prompt Generators for Adapting Language Models to CTR PredictionJianghao Lin, Bo Chen, Hangyu Wang, Yunjia Xi 等WWW 2024 · 被引用 58 次
- A Graph-Enhanced Click Model for Web SearchJianghao Lin, Weiwen Liu, Xinyi Dai, Weinan Zhang 等SIGIR 2021 · 被引用 44 次
- MAP: A Model-agnostic Pretraining Framework for Click-through Rate PredictionJianghao Lin, Yanru Qu, Wei Guo, Xinyi Dai 等KDD 2023 · 被引用 28 次
- GMOCAT: A Graph-Enhanced Multi-Objective Method for Computerized Adaptive TestingHangyu Wang, Ting Long, Liang Yin, Weinan Zhang 等KDD 2023 · 被引用 20 次
相关 Paper
- Cross-Positional Attention for Debiasing ClicksHonglei Zhuang, Zhen Qin, Xuanhui Wang, Michael Bendersky 等WWW 2021 · 被引用 43 次
- Whole Page Unbiased Learning to RankHaitao Mao, Lixin Zou, Yujia Zheng, Jiliang Tang 等WWW 2024 · 被引用 6 次
- Can Clicks Be Both Labels and Features?: Unbiased Behavior Feature Collection and Uncertainty-aware Learning to RankTao Yang, Chen Luo, Hanqing Lu, Parth Gupta 等SIGIR 2022 · 被引用 23 次
- An Offline Metric for the Debiasedness of Click ModelsRomain Deffayet, Philipp Hager, Jean-Michel Renders, Maarten de RijkeSIGIR 2023 · 被引用 7 次
- A Deep Recurrent Survival Model for Unbiased RankingJiarui Jin, Yuchen Fang, Weinan Zhang, Kan Ren 等SIGIR 2020 · 被引用 16 次
