Hierarchical Shrinkage: Improving the accuracy and interpretability of tree-based models
Abhineet Agarwal, Yan Shuo Tan, Omer Ronen, Chandan Singh, Bin Yu
Abstract
Tree-based models such as decision trees and random forests (RF) are a cornerstone of modern machine-learning practice. To mitigate overfitting, trees are typically regularized by a variety of techniques that modify their structure (e.g. pruning). We introduce Hierarchical Shrinkage (HS), a post-hoc algorithm that does not modify the tree structure, and instead regularizes the tree by shrinking the prediction over each node towards the sample means of its ancestors. The amount of shrinkage is controlled by a single regularization parameter and the number of data points in each ancestor. Since HS is a post-hoc method, it is extremely fast, compatible with any tree growing algorithm, and can be used synergistically with other regularization techniques. Extensive experiments over a wide variety of real-world datasets show that HS substantially increases the predictive performance of decision trees, even when used in conjunction with other regularization techniques. Moreover, we find that applying HS to each tree in an RF often improves accuracy, as well as its interpretability by simplifying and stabilizing its decision boundaries and SHAP values. We further explain the success of HS in improving prediction performance by showing its equivalence to ridge regression on a (supervised) basis constructed of decision stumps associated with the internal nodes of a tree. All code and models are released in a full-fledged package available on Github (github.com/csinva/imodels)
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1cc3234a-5df0-4c23-ab06-b1204d20156aCited by top-tier papers8
- "What Data Benefits My Classifier?" Enhancing Model Performance and Interpretability through Influence-Based Data SelectionAnshuman Chhabra, Peizhao Li, Prasant Mohapatra, Hongfu LiuICLR 2024 · 32 citations
- ED-Copilot: Reduce Emergency Department Wait Time with Language Model Diagnostic AssistanceLiwen Sun, Abhineet Agarwal, Aaron Kornblith, Bin Yu et al.ICML 2024 · 7 citations
- On the Convergence of CART under Sufficient Impurity Decrease ConditionRahul Mazumder, Haoyue WangNeurIPS 2023 · 7 citations
- Improving Prototypical Visual Explanations with Reward Reweighing, Reselection, and RetrainingAaron Jiaxun Li, Robin Netzorg, Zhihan Cheng, Zhuoqin Zhang et al.ICML 2024 · 5 citations
- Differentiable Decision Tree via "ReLU+Argmin" ReformulationQiangqiang Mao, Jiayang Ren, Yixiu Wang, Chenxuanyin Zou et al.NeurIPS 2025 · 2 citations
Builds on2
- Generalized and Scalable Optimal Sparse Decision TreesJimmy Lin, Chudi Zhong, Diane Hu, Cynthia Rudin et al.ICML 2020 · 174 citations
- Interpolation can hurt robust generalization even when there is no noiseKonstantin Donhauser, Alexandru Tifrea, Michael Aerni, Reinhard Heckel et al.NeurIPS 2021 · 18 citations
Related papers
- Tree Ensemble Explainability through the Hoeffding Functional Decomposition and TreeHFD AlgorithmClément BénardNeurIPS 2025 · 7 citations
- Smaller, more accurate regression forests using tree alternating optimizationArman Zharmagambetov, Miguel Á. Carreira-PerpiñánICML 2020 · 34 citations
- Smooth And Consistent Probabilistic Regression TreesSami Alkhoury, Emilie Devijver, Marianne Clausel, Myriam Tami et al.NeurIPS 2020 · 13 citations
- Beyond TreeSHAP: Efficient Computation of Any-Order Shapley Interactions for Tree EnsemblesMaximilian Muschalik, Fabian Fumagalli, Barbara Hammer, Eyke HüllermeierAAAI 2024 · 35 citations
- Linear tree shapPeng Yu, Albert Bifet, Jesse Read, Chao XuNeurIPS 2022 · 27 citations
