Invariant Random Forest: Tree-Based Model Solution for OOD Generalization
Yufan Liao, Qi Wu, Xing Yan
摘要
Out-Of-Distribution (OOD) generalization is an essential topic in machine learning. However, recent research is only focusing on the corresponding methods for neural networks. This paper introduces a novel and effective solution for OOD generalization of decision tree models, named Invariant Decision Tree (IDT). IDT enforces a penalty term with regard to the unstable/varying behavior of a split across different environments during the growth of the tree. Its ensemble version, the Invariant Random Forest (IRF), is constructed. Our proposed method is motivated by a theoretical result under mild conditions, and validated by numerical tests with both synthetic and real datasets. The superior performance compared to non-OOD tree models implies that considering OOD generalization for tree models is absolutely necessary and should be given more attention.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper9
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang 等ICML 2021 · 被引用 1,163 次
- The Risks of Invariant Risk MinimizationElan Rosenfeld, Pradeep Kumar Ravikumar, Andrej RisteskiICLR 2021 · 被引用 356 次
- Stable Prediction with Model Misspecification and Agnostic Distribution ShiftKun Kuang, Ruoxuan Xiong, Peng Cui, Susan Athey 等AAAI 2020 · 被引用 155 次
- Stable Learning via Sample ReweightingZheyan Shen, Peng Cui, Tong Zhang, Kun KuangAAAI 2020 · 被引用 155 次
- Stable Learning via Differentiated Variable DecorrelationZheyan Shen, Peng Cui, Jiashuo Liu, Tong Zhang 等KDD 2020 · 被引用 43 次
相关 Paper
- Orthogonality Matters: Invariant Time Series Representation for Out-of-distribution ClassificationRuize Shi, Hong Huang, Kehan Yin, Wei Zhou 等KDD 2024 · 被引用 5 次
- Time-Series Forecasting for Out-of-Distribution Generalization Using Invariant LearningHaoxin Liu, Harshavardhan Kamarthi, Lingkai Kong, Zhiyuan Zhao 等ICML 2024 · 被引用 32 次
- Provably Invariant Learning without Domain InformationXiaoyu Tan, Lin Yong, Shengyu Zhu, Chao Qu 等ICML 2023 · 被引用 24 次
- Heterogeneous Risk MinimizationJiashuo Liu, Zheyuan Hu, Peng Cui, Bo Li 等ICML 2021 · 被引用 170 次
- On the Connection between Invariant Learning and Adversarial Training for Out-of-Distribution GeneralizationShiji Xin, Yifei Wang, Jingtong Su, Yisen WangAAAI 2023 · 被引用 14 次
