Robust Loss Functions for Training Decision Trees with Noisy Labels
Jonathan Wilton, Nan Ye
摘要
We consider training decision trees using noisily labeled data, focusing on loss functions that can lead to robust learning algorithms. Our contributions are threefold. First, we offer novel theoretical insights on the robustness of many existing loss functions in the context of decision tree learning. We show that some of the losses belong to a class of what we call conservative losses, and the conservative losses lead to an early stopping behavior during training and noise-tolerant predictions during testing. Second, we introduce a framework for constructing robust loss functions, called distribution losses. These losses apply percentile-based penalties based on an assumed margin distribution, and they naturally allow adapting to different noise rates via a robustness parameter. In particular, we introduce a new loss called the negative exponential loss, which leads to an efficient greedy impurity-reduction learning algorithm. Lastly, our experiments on multiple datasets and noise settings validate our theoretical insight and the effectiveness of our adaptive negative exponential loss.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Using Noise to Infer Aspects of Simplicity Without LearningZachery Boner, Harry Chen, Lesia Semenova, Ronald Parr 等NeurIPS 2024 · 被引用 10 次
- Unveiling Open-set Noise: Theoretical Insights into Label NoiseChen Feng, Nicu Sebe, Georgios Tzimiropoulos, Miguel R. D. Rodrigues 等ACM MM 2025
- Identifying and Correcting Label Noise for Robust GNNs via Influence ContradictionWei Ju, Wei Zhang, Siyu Yi, Zhengyang Mao 等ICML 2026
它引用的顶会 Paper5
- Symmetric Cross Entropy for Robust Learning With Noisy LabelsYisen Wang, Xingjun Ma, Zaiyi Chen, Yuan Luo 等ICCV 2019 · 被引用 1,125 次
- Does label smoothing mitigate label noise?Michal Lukasik, Srinadh Bhojanapalli, Aditya Krishna Menon, Sanjiv KumarICML 2020 · 被引用 411 次
- NLNL: Negative Learning for Noisy LabelsYoungdong Kim, Junho Yim, Juseung Yun, Junmo KimICCV 2019 · 被引用 338 次
- Curriculum Loss: Robust Learning and Generalization against Label CorruptionYueming Lyu, Ivor W. TsangICLR 2020 · 被引用 190 次
- Positive-Unlabeled Learning using Random Forests via Recursive Greedy Risk MinimizationJonathan Wilton, Abigail M. Y. Koay, Ryan K. L. Ko, Miao Xu 等NeurIPS 2022 · 被引用 20 次
相关 Paper
- A Model-Agnostic Approach for Learning with Noisy Labels of Arbitrary DistributionsShuang Hao, Peng Li, Renzhi Wu, Xu ChuICDE 2022 · 被引用 2 次
- Popular decision tree algorithms are provably noise tolerantGuy Blanc, Jane Lange, Ali Malik, Li-Yang TanICML 2022 · 被引用 7 次
- Label Distributionally Robust Losses for Multi-class Classification: Consistency, Robustness and AdaptivityDixian Zhu, Yiming Ying, Tianbao YangICML 2023 · 被引用 15 次
- Breiman meets Bellman: Non-Greedy Decision Trees with MDPsHector Kohler, Riad Akrour, Philippe PreuxKDD 2025
- Average Sensitivity of Decision Tree LearningSatoshi Hara, Yuichi YoshidaICLR 2023
