Domain Generalization via Gradient Surgery
Lucas Mansilla, Rodrigo Echeveste, Diego H. Milone, Enzo Ferrante
Abstract
In real-life applications, machine learning models often face scenarios where there is a change in data distribution between training and test domains. When the aim is to make predictions on distributions different from those seen at training, we incur in a domain generalization problem. Methods to address this issue learn a model using data from multiple source domains, and then apply this model to the unseen target domain. Our hypothesis is that when training with multiple domains, conflicting gradients within each mini-batch contain information specific to the individual domains which is irrelevant to the others, including the test domain. If left untouched, such disagreement may degrade generalization performance. In this work, we characterize the conflicting gradients emerging in domain shift scenarios and devise novel gradient agreement strategies based on gradient surgery to alleviate their effect. We validate our approach in image classification tasks with three multi-domain datasets, showing the value of the proposed agreement strategy in enhancing the generalization capability of deep learning models in domain shift scenarios. Introduction Deep learning models have shown remarkable results in diverse application areas such as image understanding [13, 29] , speech recognition [10, 19] and natural language processing [25, 27] . Such models are typically trained under the standard supervised learning paradigm, assuming that training and test data come from the same distribution. However, in real life, training and test conditions may differ by several factors, such as a change in data acquisition device or target population. This makes models perform poorly when applied to test data whose distribution differs from the training data and, therefore, limits their implementation in such real scenarios. The goal is then to develop deep learning models that generalize outside the training distribution, under domain shift conditions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext de09fec2-5c97-4e04-8a8b-2a523327da84Cited by top-tier papers30
- Fishr: Invariant Gradient Variances for Out-of-Distribution GeneralizationAlexandre Ramé, Corentin Dancette, Matthieu CordICML 2022 · 262 citations
- Domain-General Crowd Counting in Unseen ScenariosZhipeng Du, Jiankang Deng, Miaojing ShiAAAI 2023 · 63 citations
- Towards Open-Set Test-Time Adaptation Utilizing the Wisdom of Crowds in Entropy MinimizationJungsoo Lee, Debasmit Das, Jaegul Choo, Sungha ChoiICCV 2023 · 48 citations
- Generalizable Decision Boundaries: Dualistic Meta-Learning for Open Set Domain GeneralizationXiran Wang, Jian Zhang, Lei Qi, Yinghuan ShiICCV 2023 · 39 citations
- Doodle It Yourself: Class Incremental Learning by Drawing a Few SketchesAyan Kumar Bhunia, Viswanatha Reddy Gajjala, Subhadeep Koley, Rohit Kundu et al.CVPR 2022 · 28 citations
Builds on3
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine et al.NeurIPS 2020 · 2,261 citations
- Gradient Vaccine: Investigating and Improving Multi-task Optimization in Massively Multilingual ModelsZirui Wang, Yulia Tsvetkov, Orhan Firat, Yuan CaoICLR 2021 · 241 citations
- Learning explanations that are hard to varyGiambattista Parascandolo, Alexander Neitz, Antonio Orvieto, Luigi Gresele et al.ICLR 2021 · 221 citations
Related papers
- Model-Based Domain GeneralizationAlexander Robey, George J. Pappas, Hamed HassaniNeurIPS 2021 · 167 citations
- Federated Unsupervised Domain Generalization Using Global and Local Alignment of GradientsFarhad Pourpanah, Mahdiyar Molahasani, Milad Soltany, Michael A. Greenspan et al.AAAI 2025 · 10 citations
- Symmetric Self-Paced Learning for Domain GeneralizationDi Zhao, Yun Sing Koh, Gillian Dobbie, Hongsheng Hu et al.AAAI 2024 · 16 citations
- Multi-Source Collaborative Gradient Discrepancy Minimization for Federated Domain GeneralizationYikang Wei, Yahong HanAAAI 2024 · 20 citations
- An Iterative Self-Learning Framework for Medical Domain GeneralizationZhenbang Wu, Huaxiu Yao, David M. Liebovitz, Jimeng SunNeurIPS 2023 · 12 citations
