Probabilistic Gradient Boosting Machines for Large-Scale Probabilistic Regression
Olivier Sprangers, Sebastian Schelter, Maarten de Rijke
摘要
Gradient Boosting Machines (GBMs) are hugely popular for solving tabular data problems. However, practitioners are not only interested in point predictions, but also in probabilistic predictions in order to quantify the uncertainty of the predictions. Creating such probabilistic predictions is difficult with existing GBM-based solutions: they either require training multiple models or they become too computationally expensive to be useful for large-scale settings. We propose Probabilistic Gradient Boosting Machines (PGBMs), a method to create probabilistic predictions with a single ensemble of decision trees in a computationally efficient manner. PGBM approximates the leaf weights in a decision tree as a random variable, and approximates the mean and variance of each sample in a dataset via stochastic tree ensemble update equations. These learned moments allow us to subsequently sample from a specified distribution after training. We empirically demonstrate the advantages of PGBM compared to existing state-of-the-art methods: (i) PGBM enables probabilistic estimates without compromising on point performance in a single model, (ii) PGBM learns probabilistic estimates via a single model only (and without requiring multi-parameter boosting), and thereby offers a speedup of up to several orders of magnitude over existing state-of-the-art methods on large datasets, and (iii) PGBM achieves accurate probabilistic estimates in tasks with complex differentiable loss functions, such as hierarchical time series problems, where we observed up to 10% improvement in point forecasting performance and up to 300% improvement in probabilistic forecasting performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Instance-Based Uncertainty Estimation for Gradient-Boosted Regression TreesJonathan Brophy, Daniel LowdNeurIPS 2022 · 被引用 17 次
- Probabilistic Forecasting: A Level-Set ApproachHilaf Hasson, Bernie Wang, Tim Januschowski, Jan GasthausNeurIPS 2021 · 被引用 15 次
- RashomonGB: Analyzing the Rashomon Effect and Mitigating Predictive Multiplicity in Gradient BoostingHsiang Hsu, Ivan Brugere, Shubham Sharma, Freddy Lécué 等NeurIPS 2024 · 被引用 13 次
- Treeffuser: probabilistic prediction via conditional diffusions with gradient-boosted treesNicolas Beltran-Velez, Alessandro Antonio Grande, Achille Nazaret, Alp Kucukelbir 等NeurIPS 2024 · 被引用 8 次
- CoffeeBoost: Gradient Boosting Native Conformal Inference for Bayesian OptimizationYuanhao Lai, Pengfei Zheng, Chenpeng Ji, Cheng Qiu 等AAAI 2025 · 被引用 1 次
它引用的顶会 Paper1
相关 Paper
- Uncertainty in Gradient Boosting via EnsemblesAndrey Malinin, Liudmila Prokhorenkova, Aleksei UstimenkoICLR 2021 · 被引用 117 次
- Wasserstein Gradient Boosting: A Framework for Distribution-Valued Supervised LearningTakuo MatsubaraNeurIPS 2024 · 被引用 7 次
- Neural Oblivious Decision Ensembles for Deep Learning on Tabular DataSergei Popov, Stanislav Morozov, Artem BabenkoICLR 2020 · 被引用 407 次
- SketchBoost: Fast Gradient Boosted Decision Tree for Multioutput ProblemsLeonid Iosipoi, Anton VakhrushevNeurIPS 2022 · 被引用 20 次
- Smooth And Consistent Probabilistic Regression TreesSami Alkhoury, Emilie Devijver, Marianne Clausel, Myriam Tami 等NeurIPS 2020 · 被引用 13 次
