Targeted Data-driven Regularization for Out-of-Distribution Generalization
Mohammad Mahdi Kamani, Sadegh Farhang, Mehrdad Mahdavi, James Z. Wang
Abstract
Due to biases introduced by large real-world datasets, deviations of deep learning models from their expected behavior on out-of-distribution test data are worrisome. Especially when data come from imbalanced or heavy-tailed label distributions, or minority groups of a sensitive feature. Classical approaches to address these biases are mostly data- or application-dependent, hence are burdensome to tune. Some meta-learning approaches, on the other hand, aim to learn hyperparameters in the learning process using different objective functions on training and validation data. However, these methods suffer from high computational complexity and are not scalable to large datasets. In this paper, we propose a unified data-driven regularization approach to learn a generalizable model from biased data. The proposed framework, named as targeted data-driven regularization (TDR), is model- and dataset-agnostic, and employs a target dataset that resembles the desired nature of test data in order to guide the learning process in a coupled manner. We cast the problem as a bilevel optimization and propose an efficient stochastic gradient descent based method to solve it. The framework can be utilized to alleviate various types of biases in real-world applications. We empirically show, on both synthetic and real-world datasets, the superior performance of TDR for resolving issues stem from these biases.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bbdb4827-871a-47d6-8005-6900b9018620Cited by top-tier papers3
- Self-Supervised Aggregation of Diverse Experts for Test-Agnostic Long-Tailed RecognitionYifan Zhang, Bryan Hooi, Lanqing Hong, Jiashi FengNeurIPS 2022 · 214 citations
- Robust Learning by Self-Transition for Handling Noisy LabelsHwanjun Song, Minseok Kim, Dongmin Park, Yooju Shin et al.KDD 2021 · 27 citations
- Nonlinearity Encoding for Extrapolation of Neural NetworksGyoung S. Na, Chanyoung ParkKDD 2022 · 4 citations
Builds on2
- Test-Time Training with Self-Supervision for Generalization under Distribution ShiftsYu Sun, Xiaolong Wang, Zhuang Liu, John Miller et al.ICML 2020 · 1,220 citations
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang et al.ICML 2021 · 1,163 citations
Related papers
- Heteroskedastic and Imbalanced Deep Learning with Adaptive RegularizationKaidi Cao, Yining Chen, Junwei Lu, Nikos Aréchiga et al.ICLR 2021 · 20 citations
- Meta-Learning with Fewer Tasks through Task InterpolationHuaxiu Yao, Linjun Zhang, Chelsea FinnICLR 2022 · 66 citations
- Regularizing Neural Networks with Meta-Learning Generative ModelsShin'ya Yamaguchi, Daiki Chijiwa, Sekitoshi Kanai, Atsutoshi Kumagai et al.NeurIPS 2023 · 10 citations
- Fair Bilevel Neural Network (FairBiNN): On Balancing fairness and accuracy via Stackelberg EquilibriumMehdi Yazdani-Jahromi, Ali Khodabandeh Yalabadi, Amirarsalan Rajabi, Aida Tayebi et al.NeurIPS 2024 · 11 citations
- Meta-Learning without MemorizationMingzhang Yin, George Tucker, Mingyuan Zhou, Sergey Levine et al.ICLR 2020 · 201 citations
