Adaptive Risk Minimization: Learning to Adapt to Domain Shift
Marvin Zhang, Henrik Marklund, Nikita Dhawan, Abhishek Gupta, Sergey Levine, Chelsea Finn
Abstract
A fundamental assumption of most machine learning algorithms is that the training and test data are drawn from the same underlying distribution. However, this assumption is violated in almost all practical applications: machine learning systems are regularly tested under distribution shift, due to changing temporal correlations, atypical end users, or other factors. In this work, we consider the problem setting of domain generalization, where the training data are structured into domains and there may be multiple test time shifts, corresponding to new domains or domain distributions. Most prior methods aim to learn a single robust model or invariant feature space that performs well on all domains. In contrast, we aim to learn models that adapt at test time to domain shift using unlabeled test points. Our primary contribution is to introduce the framework of adaptive risk minimization (ARM), in which models are directly optimized for effective adaptation to shift by learning to adapt on the training domains. Compared to prior methods for robustness, invariance, and adaptation, ARM methods provide performance gains of 1-4% test accuracy on a number of image classification problems exhibiting domain shift.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1df19148-bdfa-450d-bbfc-e556142fb296Cited by top-tier papers75
- MEMO: Test Time Robustness via Adaptation and AugmentationMarvin Zhang, Sergey Levine, Chelsea FinnNeurIPS 2022 · 595 citations
- The Norm Must Go On: Dynamic Unsupervised Domain Adaptation by NormalizationMuhammad Jehanzeb Mirza, Jakub Micorek, Horst Possegger, Horst BischofCVPR 2022 · 119 citations
- OST: Improving Generalization of DeepFake Detection via One-Shot Test-Time TrainingLiang Chen, Yong Zhang, Yibing Song, Jue Wang et al.NeurIPS 2022 · 102 citations
- MADG: Margin-based Adversarial Learning for Domain GeneralizationAveen Dayal, Vimal K. B., Linga Reddy Cenkeramaddi, C. Krishna Mohan et al.NeurIPS 2023 · 102 citations
- Probable Domain Generalization via Quantile Risk MinimizationCian Eastwood, Alexander Robey, Shashank Singh, Julius von Kügelgen et al.NeurIPS 2022 · 99 citations
Builds on9
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang et al.ICCV 2019 · 2,239 citations
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- Tent: Fully Test-Time Adaptation by Entropy MinimizationDequan Wang, Evan Shelhamer, Shaoteng Liu, Bruno A. Olshausen et al.ICLR 2021 · 1,731 citations
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 1,416 citations
- Test-Time Training with Self-Supervision for Generalization under Distribution ShiftsYu Sun, Xiaolong Wang, Zhuang Liu, John Miller et al.ICML 2020 · 1,220 citations
Related papers
- Learning Optimal Features via Partial InvarianceMoulik Choraria, Ibtihal Ferwana, Ankur Mani, Lav R. VarshneyAAAI 2023 · 3 citations
- Domain-Specific Risk Minimization for Domain GeneralizationYifan Zhang, Jindong Wang, Jian Liang, Zhang Zhang et al.KDD 2023 · 26 citations
- Bayesian Adaptation for Covariate ShiftAurick Zhou, Sergey LevineNeurIPS 2021 · 40 citations
- Distribution Shift Is Key to Learning Invariant PredictionHong Zheng, Fei TengAAAI 2026
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang et al.ICML 2021 · 1,163 citations
