MA4DIV: Multi-Agent Reinforcement Learning for Search Result Diversification
Yiqun Chen, Jiaxin Mao, Yi Zhang, Dehong Ma, Long Xia, Jun Fan, Daiting Shi, Zhicong Cheng, Simiu Gu, Dawei Yin
Abstract
Search result diversification (SRD), which aims to ensure that documents in a ranking list cover a broad range of subtopics, is a significant and widely studied problem in Information Retrieval and Web Search. Existing methods primarily utilize a paradigm of "greedy selection", i.e., selecting one document with the highest diversity score at a time or optimize an approximation of the objective function. These approaches tend to be inefficient and are easily trapped in a suboptimal state. To address these challenges, we introduce Multi-Agent reinforcement learning (MARL) for search result DIVersity, which called MA4DIV 1 . In this approach, each document is an agent and the search result diversification is modeled as a cooperative task among multiple agents. By modeling the SRD ranking problem as a cooperative MARL problem, this approach allows for directly optimizing the diversity metrics, such as 𝛼-NDCG, while achieving high training efficiency. We conducted experiments on public TREC datasets and a larger scale dataset in the industrial setting. The experiemnts show that MA4DIV achieves substantial improvements in both effectiveness and efficiency than existing baselines, especially on the industrial dataset. CCS Concepts • Information systems → Information retrieval diversity. * Jiaxin Mao and Dawei Yin are the corresponding authors. 1 The code of MA4DIV can be seen on https://github.com/chenyiqun/MA4DIV .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0786d0cc-8bb2-4a5c-92e6-2dce77fc22ecBuilds on8
- QPLEX: Duplex Dueling Multi-Agent Q-LearningJianhao Wang, Zhizhou Ren, Terry Liu, Yang Yu et al.ICLR 2021 · 595 citations
- SetRank: Learning a Permutation-Invariant Ranking Model for Information RetrievalLiang Pang, Jun Xu, Qingyao Ai, Yanyan Lan et al.SIGIR 2020 · 113 citations
- Diversification-Aware Learning to Rank using Distributed RepresentationLe Yan, Zhen Qin, Rama Kumar Pasumarthi, Xuanhui Wang et al.WWW 2021 · 44 citations
- RLPer: A Reinforcement Learning Model for Personalized SearchJing Yao, Zhicheng Dou, Jun Xu, Ji-Rong WenWWW 2020 · 33 citations
- Modeling Intent Graph for Search Result DiversificationZhan Su, Zhicheng Dou, Yutao Zhu, Xubo Qin et al.SIGIR 2021 · 32 citations
Related papers
- Optimize What You Evaluate With: Search Result Diversification Based on Metric OptimizationHai-Tao YuAAAI 2022 · 11 citations
- Reinforcement Learning to Rank with Pairwise Policy GradientJun Xu, Zeng Wei, Long Xia, Yanyan Lan et al.SIGIR 2020 · 32 citations
- Controlling Behavioral Diversity in Multi-Agent Reinforcement LearningMatteo Bettini, Ryan Kortvelesy, Amanda ProrokICML 2024 · 11 citations
- Can Cooperative Multi-Agent Reinforcement Learning Boost Automatic Web Testing? An Exploratory StudyYujia Fan, Sinan Wang, Zebang Fei, Yao Qin et al.ASE 2024 · 3 citations
- MARS²: Scaling Multi-Agent Tree Search via Reinforcement Learning for Code GenerationPengfei Li, Shijie Wang, Fangyuan Li, Yikun Fu et al.ACL 2026
