Diversification-Aware Learning to Rank using Distributed Representation
Le Yan, Zhen Qin, Rama Kumar Pasumarthi, Xuanhui Wang, Michael Bendersky
Abstract
Existing work on search result diversification typically falls into the “next document” paradigm, that is, selecting the next document based on the ones already chosen. A sequential process of selecting documents one-by-one is naturally modeled in learning-based approaches. However, such a process makes the learning difficult because there are an exponential number of ranking lists to consider. Sampling is usually used to reduce the computational complexity but this makes the learning less effective. In this paper, we propose a soft version of the “next document” paradigm in which we associate each document with an approximate rank, and thus the subtopics covered prior to a document can also be estimated. We show that we can derive differentiable diversification-aware losses, which are smooth approximation of diversity metrics like α-NDCG, based on these estimates. We further propose to optimize the losses in the learning-to-rank setting using neural distributed representations of queries and documents. Experiments are conducted on the public benchmark TREC datasets. By comparing with an extensive list of baseline methods, we show that our Diversification-Aware LEarning-TO-Rank (DALETOR) approaches outperform them by a large margin, while being much simpler during learning and inference.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f26d74e8-5dca-4436-8aad-54aef6127b6eCited by top-tier papers5
- LLM4Rerank: LLM-based Auto-Reranking Framework for RecommendationsJingtong Gao, Bo Chen, Xiangyu Zhao, Weiwen Liu et al.WWW 2025 · 50 citations
- Towards Explainable Search Results: A Listwise Explanation GeneratorPuxuan Yu, Razieh Rahimi, James AllanSIGIR 2022 · 26 citations
- Knowledge Enhanced Search Result DiversificationZhan Su, Zhicheng Dou, Yutao Zhu, Ji-Rong WenKDD 2022 · 16 citations
- Optimize What You Evaluate With: Search Result Diversification Based on Metric OptimizationHai-Tao YuAAAI 2022 · 11 citations
- MA4DIV: Multi-Agent Reinforcement Learning for Search Result DiversificationYiqun Chen, Jiaxin Mao, Yi Zhang, Dehong Ma et al.WWW 2025 · 7 citations
Builds on3
- SetRank: Learning a Permutation-Invariant Ranking Model for Information RetrievalLiang Pang, Jun Xu, Qingyao Ai, Yanyan Lan et al.SIGIR 2020 · 113 citations
- Are Neural Rankers still Outperformed by Gradient Boosted Decision Trees?Zhen Qin, Le Yan, Honglei Zhuang, Yi Tay et al.ICLR 2021 · 41 citations
- DVGAN: A Minimax Game for Search Result Diversification Combining Explicit and Implicit FeaturesJiongnan Liu, Zhicheng Dou, Xiaojie Wang, Shuqi Lu et al.SIGIR 2020 · 32 citations
Related papers
- A Guided Learning Approach for Item Recommendation via Surrogate Loss LearningAhmed Rashed, Josif Grabocka, Lars Schmidt-ThiemeSIGIR 2021 · 10 citations
- Modeling Intent Graph for Search Result DiversificationZhan Su, Zhicheng Dou, Yutao Zhu, Xubo Qin et al.SIGIR 2021 · 32 citations
- An Alternative Cross Entropy Loss for Learning-to-RankSebastian BruchWWW 2021 · 58 citations
- A Reinforcement Learning Framework for Relevance FeedbackAli Montazeralghaem, Hamed Zamani, James AllanSIGIR 2020 · 38 citations
- Efficient Neural Ranking using Forward IndexesJurek Leonhardt, Koustav Rudra, Megha Khosla, Abhijit Anand et al.WWW 2022 · 16 citations
