Bivariate Decision Trees: Smaller, Interpretable, More Accurate
Rasul Kairgeldin, Miguel Á. Carreira-Perpiñán
Abstract
Univariate decision trees, commonly used since the 1950s, predict by asking questions about a single feature in each decision node. While they are interpretable, they often lack competitive predictive accuracy due to their inability to model feature correlations. Multivariate (oblique) trees use multiple features in each node, capturing high-dimensional correlations better, but sometimes they can be difficult to interpret. We advocate for a model that strikes a useful middle ground: bivariate decision trees, which use two features in each node. This typically produces trees that not only are more accurate than univariate trees, but much smaller, which offsets the small increase in node complexity and keeps them interpretable. They also help data mining by constructing new features that are useful for discrimination, and by providing a form of supervised, hierarchical 2D visualization that reveals patterns such as clusters or linear structure. We give two new algorithms to learn bivariate trees: a fast one based on CART; and a slower one based on alternating optimization with a feature regularization term, which produces the best trees while still scaling to large datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 541ddc9a-2f3d-4050-a124-78d5894a4f48Cited by top-tier papers3
- Exact Functional ANOVA Decomposition for Categorical InputsBaptiste Ferrere, Nicolas Bousquet, Gamboa Fabrice, Jean-Michel Loubes et al.ICML 2026 · 1 citation
- Breiman meets Bellman: Non-Greedy Decision Trees with MDPsHector Kohler, Riad Akrour, Philippe PreuxKDD 2025
- Empowering Decision Trees via Shape Function BranchingNakul Upadhya, Eldan CohenNeurIPS 2025
Builds on7
- Counterfactual Explanations for Oblique Decision Trees: Exact, Efficient AlgorithmsMiguel Á. Carreira-Perpiñán, Suryabhan Singh HadaAAAI 2021 · 39 citations
- Smaller, more accurate regression forests using tree alternating optimizationArman Zharmagambetov, Miguel Á. Carreira-PerpiñánICML 2020 · 34 citations
- Optimal Interpretable Clustering Using Oblique Decision TreesMagzhan Gabidolla, Miguel Á. Carreira-PerpiñánKDD 2022 · 16 citations
- Pushing the Envelope of Gradient Boosting Forests via Globally-Optimized Oblique TreesMagzhan Gabidolla, Miguel Á. Carreira-PerpiñánCVPR 2022 · 13 citations
- Beyond the ROC Curve: Classification Trees Using Cost-Optimal Curves, with Application to Imbalanced DatasetsMagzhan Gabidolla, Arman Zharmagambetov, Miguel Á. Carreira-PerpiñánICML 2024 · 5 citations
Related papers
- Explanations of Black-Box Models based on Directional Feature InteractionsAria Masoomi, Davin Hill, Zhonghui Xu, Craig P. Hersh et al.ICLR 2022 · 26 citations
- Sparse Learning with CARTJason M. KlusowskiNeurIPS 2020 · 31 citations
- Feature Learning for Interpretable, Performant Decision TreesJack H. Good, Torin Kovach, Kyle Miller, Artur DubrawskiNeurIPS 2023 · 16 citations
- Explainable k-Means and k-Medians ClusteringMichal Moshkovitz, Sanjoy Dasgupta, Cyrus Rashtchian, Nave FrostICML 2020 · 184 citations
- Decision trees as partitioning machines to characterize their generalization propertiesJean-Samuel Leboeuf, Frédéric Leblanc, Mario MarchandNeurIPS 2020 · 17 citations
