Are Neural Rankers still Outperformed by Gradient Boosted Decision Trees?
Zhen Qin, Le Yan, Honglei Zhuang, Yi Tay, Rama Kumar Pasumarthi, Xuanhui Wang, Michael Bendersky, Marc Najork
摘要
Despite the success of neural models on many major machine learning problems, their effectiveness on traditional Learning-to-Rank (LTR) problems is still not widely acknowledged. We first validate this concern by showing that most recent neural LTR models are, by a large margin, inferior to the best publicly available Gradient Boosted Decision Trees (GBDT) in terms of their reported ranking accuracy on benchmark datasets. This unfortunately was somehow overlooked in recent neural LTR papers. We then investigate why existing neural LTR models under-perform and identify several of their weaknesses. Furthermore, we propose a unified framework comprising of counter strategies to ameliorate the existing weaknesses of neural models. Our models are the first to be able to perform equally well, comparing with the best tree-based baseline, while outperforming recently published neural LTR models by a large margin. Our results can also serve as a benchmark to facilitate future improvement of neural LTR models.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper25
- Revisiting Deep Learning Models for Tabular DataYury Gorishniy, Ivan Rubachev, Valentin Khrulkov, Artem BabenkoNeurIPS 2021 · 被引用 1,847 次
- Quantized Training of Gradient Boosting Decision TreesYu Shi, Guolin Ke, Zhuoming Chen, Shuxin Zheng 等NeurIPS 2022 · 被引用 51 次
- Diversification-Aware Learning to Rank using Distributed RepresentationLe Yan, Zhen Qin, Rama Kumar Pasumarthi, Xuanhui Wang 等WWW 2021 · 被引用 44 次
- Cross-Positional Attention for Debiasing ClicksHonglei Zhuang, Zhen Qin, Xuanhui Wang, Michael Bendersky 等WWW 2021 · 被引用 43 次
- Toward Understanding Privileged Features Distillation in Learning-to-RankShuo Yang, Sujay Sanghavi, Holakou Rahmanian, Jan Bakus 等NeurIPS 2022 · 被引用 31 次
相关 Paper
- Which Tricks are Important for Learning to Rank?Ivan Lyzhin, Aleksei Ustimenko, Andrey Gulin, Liudmila ProkhorenkovaICML 2023 · 被引用 8 次
- Fast Attention-based Learning-To-Rank Model for Structured Map SearchChiqun Zhang, Michael R. Evans, Max Lepikhin, Dragomir YankovSIGIR 2021 · 被引用 4 次
- StochasticRank: Global Optimization of Scale-Free Discrete FunctionsAleksei Ustimenko, Liudmila ProkhorenkovaICML 2020 · 被引用 21 次
- iLTM: Integrated Large Tabular ModelDavid Bonet, Marçal Comajoan Cara, Alvaro Calafell, Daniel Mas Montserrat 等KDD 2026 · 被引用 4 次
- OPS: An Order-Preserving Sorting Network for Information RetrievalChao Wang, Yongxiang Tang, Guikai Luan, Kaiyuan Li 等SIGIR 2026
