Distributed Task-Based Training of Tree Models
Da Yan, Md Mashiur Rahman Chowdhury, Guimu Guo, Jalal Khalil, Zhe Jiang, Sushil K. Prasad
Abstract
Decision trees and tree ensembles are popular supervised learning models on tabular data. Two recent research trends on tree models stand out: (1) bigger and deeper models with many trees, and (2) scalable distributed training frameworks. However, existing implementations on distributed systems are IO-bound leaving CPU cores underutilized. They also only find best node-splitting conditions approximately due to row-based data partitioning scheme. In this paper, we target the exact training of tree models by effectively utilizing the available CPU cores. The resulting system called TreeServer adopts a column-based data partitioning scheme to minimize communication, and a node-centric task-based engine to fully explore the CPU parallelism. Experiments show that TreeServer is up to 10× faster than models in Spark MLlib. We also showcase TreeServer's high training throughput by using it to build big “deep forest” models.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Related papers
- Neural Oblivious Decision Ensembles for Deep Learning on Tabular DataSergei Popov, Stanislav Morozov, Artem BabenkoICLR 2020 · 407 citations
- Treebeard: An Optimizing Compiler for Decision Tree Based ML InferenceAshwin Prasad, Sampath Rajendra, Kaushik Rajan, R. Govindarajan et al.MICRO 2022 · 8 citations
- GRANDE: Gradient-Based Decision Tree Ensembles for Tabular DataSascha Marton, Stefan Lüdtke, Christian Bartelt, Heiner StuckenschmidtICLR 2024 · 13 citations
- C olumnSGD: A Column-oriented Framework for Distributed Stochastic Gradient DescentZhipeng Zhang, Wentao Wu, Jiawei Jiang, Lele Yu et al.ICDE 2020 · 6 citations
- Communication Algorithm-Architecture Co-Design for Distributed Deep LearningJiayi Huang, Pritam Majumder, Sungkeun Kim, Abdullah Muzahid et al.ISCA 2021 · 44 citations
