Subgroup Robustness Grows On Trees: An Empirical Baseline Investigation
Josh Gardner, Zoran Popovic, Ludwig Schmidt
摘要
Researchers have proposed many methods for fair and robust machine learning, but comprehensive empirical evaluation of their subgroup robustness is lacking. In this work, we address this gap in the context of tabular data, where sensitive subgroups are clearly-defined, real-world fairness problems abound, and prior works often do not compare to state-of-the-art tree-based models as baselines. We conduct an empirical comparison of several previously-proposed methods for fair and robust learning alongside state-of-the-art tree-based methods and other baselines. Via experiments with more than model configurations on eight datasets, we show that tree-based methods have strong subgroup robustness, even when compared to robustness- and fairness-enhancing methods. Moreover, the best tree-based models tend to show good performance over a range of metrics, while robust or group-fair models can show brittleness, with significant performance differences across different metrics for a fixed model. We also demonstrate that tree-based models show less sensitivity to hyperparameter configurations, and are less costly to train. Our work suggests that tree-based ensemble models make an effective baseline for tabular data, and are a sensible default when subgroup robustness is desired. For associated code and detailed results, see https://github.com/jpgard/subgroup-robustness-grows-on-trees .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Large Scale Transfer Learning for Tabular Data via Language ModelingJosh Gardner, Juan C. Perdomo, Ludwig SchmidtNeurIPS 2024 · 被引用 103 次
- CARTE: Pretraining and Transfer for Tabular LearningMyung Jun Kim, Léo Grinsztajn, Gaël VaroquauxICML 2024 · 被引用 52 次
- FARE: Provably Fair Representation Learning with Practical CertificatesNikola Jovanovic, Mislav Balunovic, Dimitar Iliev Dimitrov, Martin T. VechevICML 2023 · 被引用 21 次
- Multi-group Learning for Hierarchical GroupsSamuel Deng, Daniel HsuICML 2024 · 被引用 7 次
- When do Minimax-fair Learning and Empirical Risk Minimization Coincide?Harvineet Singh, Matthäus Kleindessner, Volkan Cevher, Rumi Chunara 等ICML 2023 · 被引用 6 次
它引用的顶会 Paper17
- The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution GeneralizationDan Hendrycks, Steven Basart, Norman Mu, Saurav Kadavath 等ICCV 2021 · 被引用 2,294 次
- TabNet: Attentive Interpretable Tabular LearningSercan Ö. Arik, Tomas PfisterAAAI 2021 · 被引用 2,148 次
- Revisiting Deep Learning Models for Tabular DataYury Gorishniy, Ivan Rubachev, Valentin Khrulkov, Artem BabenkoNeurIPS 2021 · 被引用 1,847 次
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 被引用 1,578 次
相关 Paper
- Individually Fair Gradient BoostingAlexander Vargo, Fan Zhang, Mikhail Yurochkin, Yuekai SunICLR 2021 · 被引用 16 次
- Sensitivity Verification for Additive Decision Tree EnsemblesArhaan Ahmad, Tanay Vineet Tayal, Ashutosh Gupta, S. AkshayICLR 2025
- Towards Understanding Fairness and its Composition in Ensemble Machine LearningUsman Gohar, Sumon Biswas, Hridesh RajanICSE 2023 · 被引用 30 次
- Robust Counterfactual Explanations for Tree-Based EnsemblesSanghamitra Dutta, Jason Long, Saumitra Mishra, Cecilia Tilli 等ICML 2022 · 被引用 73 次
- Data-Aware and Scalable Sensitivity Analysis for Decision Tree EnsemblesNamrita Varshney, Ashutosh Gupta, Arhaan Ahmad, Tanay Vineet Tayal 等ICLR 2026 · 被引用 2 次
