A Neural Tangent Kernel Perspective of Infinite Tree Ensembles
Ryuichi Kanoh, Mahito Sugiyama
Abstract
In practical situations, the tree ensemble is one of the most popular models along with neural networks. A soft tree is a variant of a decision tree. Instead of using a greedy method for searching splitting rules, the soft tree is trained using a gradient method in which the entire splitting operation is formulated in a differentiable form. Although ensembles of such soft trees have been used increasingly in recent years, little theoretical work has been done to understand their behavior. By considering an ensemble of infinite soft trees, this paper introduces and studies the Tree Neural Tangent Kernel (TNTK), which provides new insights into the behavior of the infinite ensemble of soft trees. Using the TNTK, we theoretically identify several non-trivial properties, such as global convergence of the training, the equivalence of the oblivious tree structure, and the degeneracy of the TNTK induced by the deepening of the trees.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b2d1e353-7f5c-4c70-9857-8bb2e1d61d36Cited by top-tier papers5
- Gradient Boosting Performs Gaussian Process InferenceAleksei Ustimenko, Artem Beliakov, Liudmila ProkhorenkovaICLR 2023 · 3 citations
- No-Regret Bandit Exploration based on Soft Tree Ensemble ModelShogo Iwazaki, Shinya SuzumuraNeurIPS 2024 · 3 citations
- Linear Mode Connectivity in Differentiable Tree EnsemblesRyuichi Kanoh, Mahito SugiyamaICLR 2025
- Analyzing Tree Architectures in Ensembles via Neural Tangent KernelRyuichi Kanoh, Mahito SugiyamaICLR 2023
- Neural Tangent Kernels for Axis-Aligned Tree EnsemblesRyuichi Kanoh, Mahito SugiyamaICML 2024
Builds on9
- TabNet: Attentive Interpretable Tabular LearningSercan Ö. Arik, Tomas PfisterAAAI 2021 · 2,148 citations
- GShard: Scaling Giant Models with Conditional Computation and Automatic ShardingDmitry Lepikhin, HyoukJoong Lee, Yuanzhong Xu, Dehao Chen et al.ICLR 2021 · 1,954 citations
- Neural Oblivious Decision Ensembles for Deep Learning on Tabular DataSergei Popov, Stanislav Morozov, Artem BabenkoICLR 2020 · 407 citations
- Harnessing the Power of Infinitely Wide Deep Nets on Small-data TasksSanjeev Arora, Simon S. Du, Zhiyuan Li, Ruslan Salakhutdinov et al.ICLR 2020 · 167 citations
- Why Do Deep Residual Networks Generalize Better than Deep Feedforward Networks? - A Neural Tangent Kernel PerspectiveKaixuan Huang, Yuqing Wang, Molei Tao, Tuo ZhaoNeurIPS 2020 · 107 citations
Related papers
- The Tree Ensemble Layer: Differentiability meets Conditional ComputationHussein Hazimeh, Natalia Ponomareva, Petros Mol, Zhenyu Tan et al.ICML 2020 · 95 citations
- Dynamics of Deep Neural Networks and Neural Tangent HierarchyJiaoyang Huang, Horng-Tzer YauICML 2020 · 167 citations
- FL-NTK: A Neural Tangent Kernel-based Framework for Federated Learning AnalysisBaihe Huang, Xiaoxiao Li, Zhao Song, Xin YangICML 2021 · 66 citations
- Collegial EnsemblesEtai Littwin, Ben Myara, Sima Sabah, Joshua M. Susskind et al.NeurIPS 2020 · 10 citations
- Flexible Modeling and Multitask Learning using Differentiable Tree EnsemblesShibal Ibrahim, Hussein Hazimeh, Rahul MazumderKDD 2022 · 3 citations
