A Neural Tangent Kernel Perspective of Infinite Tree Ensembles
Ryuichi Kanoh, Mahito Sugiyama
摘要
In practical situations, the tree ensemble is one of the most popular models along with neural networks. A soft tree is a variant of a decision tree. Instead of using a greedy method for searching splitting rules, the soft tree is trained using a gradient method in which the entire splitting operation is formulated in a differentiable form. Although ensembles of such soft trees have been used increasingly in recent years, little theoretical work has been done to understand their behavior. By considering an ensemble of infinite soft trees, this paper introduces and studies the Tree Neural Tangent Kernel (TNTK), which provides new insights into the behavior of the infinite ensemble of soft trees. Using the TNTK, we theoretically identify several non-trivial properties, such as global convergence of the training, the equivalence of the oblivious tree structure, and the degeneracy of the TNTK induced by the deepening of the trees.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Gradient Boosting Performs Gaussian Process InferenceAleksei Ustimenko, Artem Beliakov, Liudmila ProkhorenkovaICLR 2023 · 被引用 3 次
- No-Regret Bandit Exploration based on Soft Tree Ensemble ModelShogo Iwazaki, Shinya SuzumuraNeurIPS 2024 · 被引用 3 次
- Linear Mode Connectivity in Differentiable Tree EnsemblesRyuichi Kanoh, Mahito SugiyamaICLR 2025
- Analyzing Tree Architectures in Ensembles via Neural Tangent KernelRyuichi Kanoh, Mahito SugiyamaICLR 2023
- Neural Tangent Kernels for Axis-Aligned Tree EnsemblesRyuichi Kanoh, Mahito SugiyamaICML 2024
它引用的顶会 Paper9
- TabNet: Attentive Interpretable Tabular LearningSercan Ö. Arik, Tomas PfisterAAAI 2021 · 被引用 2,148 次
- GShard: Scaling Giant Models with Conditional Computation and Automatic ShardingDmitry Lepikhin, HyoukJoong Lee, Yuanzhong Xu, Dehao Chen 等ICLR 2021 · 被引用 1,954 次
- Neural Oblivious Decision Ensembles for Deep Learning on Tabular DataSergei Popov, Stanislav Morozov, Artem BabenkoICLR 2020 · 被引用 407 次
- Harnessing the Power of Infinitely Wide Deep Nets on Small-data TasksSanjeev Arora, Simon S. Du, Zhiyuan Li, Ruslan Salakhutdinov 等ICLR 2020 · 被引用 167 次
- Why Do Deep Residual Networks Generalize Better than Deep Feedforward Networks? - A Neural Tangent Kernel PerspectiveKaixuan Huang, Yuqing Wang, Molei Tao, Tuo ZhaoNeurIPS 2020 · 被引用 107 次
相关 Paper
- The Tree Ensemble Layer: Differentiability meets Conditional ComputationHussein Hazimeh, Natalia Ponomareva, Petros Mol, Zhenyu Tan 等ICML 2020 · 被引用 95 次
- Dynamics of Deep Neural Networks and Neural Tangent HierarchyJiaoyang Huang, Horng-Tzer YauICML 2020 · 被引用 167 次
- FL-NTK: A Neural Tangent Kernel-based Framework for Federated Learning AnalysisBaihe Huang, Xiaoxiao Li, Zhao Song, Xin YangICML 2021 · 被引用 66 次
- Collegial EnsemblesEtai Littwin, Ben Myara, Sima Sabah, Joshua M. Susskind 等NeurIPS 2020 · 被引用 10 次
- Flexible Modeling and Multitask Learning using Differentiable Tree EnsemblesShibal Ibrahim, Hussein Hazimeh, Rahul MazumderKDD 2022 · 被引用 3 次
