PINE: Pruning Boosted Tree Ensembles with Conformal In-Distribution Prediction Equivalence
Haruki Yajima, Yusuke Matsui
Abstract
Tree ensembles are machine learning models with strong predictive performance and interpretability, and remain widely used for tabular data. Standard pruning methods for tree ensembles typically optimize an accuracy-compression trade-off and may change a subset of predictions, potentially compromising decision consistency. Faithful pruning methods address this issue by preserving prediction equivalence over the entire input space, but this requirement leads to lower compression ratios. We propose PINE, a pruning method that provides strong guarantees within an in-distribution region. PINE preserves prediction equivalence within this region and controls the region size using a single parameter via conformal calibration. Experiments on 12 public tabular datasets show that PINE improves the compression ratio by up to 30% while preserving predictions at a comparable level to existing faithful pruning methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0979c118-2bb7-46f1-9a4f-057650712d12Builds on7
- Adaptive Conformal Inference Under Distribution ShiftIsaac Gibbs, Emmanuel J. CandèsNeurIPS 2021 · 665 citations
- Optimal Counterfactual Explanations in Tree EnsemblesAxel Parmentier, Thibaut VidalICML 2021 · 66 citations
- Born-Again Tree EnsemblesThibaut Vidal, Maximilian SchifferICML 2020 · 62 citations
- Versatile Verification of Tree EnsemblesLaurens Devos, Wannes Meert, Jesse DavisICML 2021 · 16 citations
- Locally Adaptive Label Smoothing Improves Predictive ChurnDara Bahri, Heinrich JiangICML 2021 · 16 citations
Related papers
- Free Lunch in the Forest: Functionally-Identical Pruning of Boosted Tree EnsemblesYoussouf Emine, Alexandre Forel, Idriss Malek, Thibaut VidalAAAI 2025 · 3 citations
- Compressing tree ensembles through Level-wise Optimization and PruningLaurens Devos, Timo Martens, Deniz Can Oruc, Wannes Meert et al.ICML 2025
- Smooth And Consistent Probabilistic Regression TreesSami Alkhoury, Emilie Devijver, Marianne Clausel, Myriam Tami et al.NeurIPS 2020 · 13 citations
- Subgroup Robustness Grows On Trees: An Empirical Baseline InvestigationJosh Gardner, Zoran Popovic, Ludwig SchmidtNeurIPS 2022 · 27 citations
- MEPSI: An MDL-Based Ensemble Pruning Approach with Structural InformationXiao-Dong Bi, Shao-Qun Zhang, Yuan JiangAAAI 2024 · 2 citations
