Oblique Decision Trees from Derivatives of ReLU Networks
Guang-He Lee, Tommi S. Jaakkola
Abstract
We show how neural models can be used to realize piece-wise constant functions such as decision trees. The proposed architecture, which we call locally constant networks, builds on ReLU networks that are piece-wise linear and hence their associated gradients with respect to the inputs are locally constant. We formally establish the equivalence between the classes of locally constant networks and decision trees. Moreover, we highlight several advantageous properties of locally constant networks, including how they realize decision trees with parameter sharing across branching / leaves. Indeed, only neurons suffice to implicitly model an oblique decision tree with leaf nodes. The neural representation also enables us to adopt many tools developed for deep networks (e.g., DropConnect (Wan et al., 2013)) while implicitly training decision trees. We demonstrate that our method outperforms alternative techniques for training oblique decision trees in the context of molecular property classification and regression tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0f5ebb8f-5e34-40c5-9f81-a12cfa5956c1Cited by top-tier papers5
- Learning Prescriptive ReLU NetworksWei Sun, Asterios TsiourvasICML 2023 · 3 citations
- Convex Polytope Trees and its Application to VAEMohammadreza Armandpour, Ali Sadeghian, Mingyuan ZhouNeurIPS 2021 · 2 citations
- Differentiable Decision Tree via "ReLU+Argmin" ReformulationQiangqiang Mao, Jiayang Ren, Yixiu Wang, Chenxuanyin Zou et al.NeurIPS 2025 · 2 citations
- Unveiling Options with Neural Network DecompositionMahdi Alikhasi, Levi LelisICLR 2024 · 2 citations
- Deep Networks Learn Features From Local Discontinuities in the Label FunctionPrithaj Banerjee, Harish Guruprasad Ramaswamy, Mahesh Lorik Yadav, Chandra Shekar LakshminarayananICLR 2025
Related papers
- Learning Binary Decision Trees by Argmin DifferentiationValentina Zantedeschi, Matt J. Kusner, Vlad NiculaeICML 2021 · 16 citations
- Deep Molecular Programming: A Natural Implementation of Binary-Weight ReLU Neural NetworksMarko Vasic, Cameron T. Chalk, Sarfraz Khurshid, David SoloveichikICML 2020 · 17 citations
- Towards Lower Bounds on the Depth of ReLU Neural NetworksChristoph Hertrich, Amitabh Basu, Marco Di Summa, Martin SkutellaNeurIPS 2021 · 70 citations
- Improved Bounds on Neural Complexity for Representing Piecewise Linear FunctionsKuan-Lin Chen, Harinath Garudadri, Bhaskar D. RaoNeurIPS 2022 · 36 citations
- Discrete Tree Flows via Tree-Structured PermutationsMai Elkady, Hyung Zin Lim, David I. InouyeICML 2022 · 2 citations
