Domain Generalization Guided by Gradient Signal to Noise Ratio of Parameters
Mateusz Michalkiewicz, Masoud Faraki, Xiang Yu, Manmohan Chandraker, Mahsa Baktashmotlagh
摘要
Overfitting to the source domain is a common issue in gradient-based training of deep neural networks. To compensate for the over-parameterized models, numerous regularization techniques have been introduced such as those based on dropout. While these methods achieve significant improvements on classical benchmarks such as ImageNet, their performance diminishes with the introduction of domain shift in the test set i.e. when the unseen data comes from a significantly different distribution. In this paper, we move away from the classical approach of Bernoulli sampled dropout mask construction and propose to base the selection on gradient-signal-to-noise ratio (GSNR) of network’s parameters. Specifically, at each training step, parameters with high GSNR will be discarded. Furthermore, we alleviate the burden of manually searching for the optimal dropout ratio by leveraging a meta-learning approach. We evaluate our method on standard domain generalization benchmarks and achieve competitive results on classification and face anti-spoofing problems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- A Layer Selection Approach to Test Time AdaptationSabyasachi Sahoo, Mostafa ElAraby, Jonas Ngnawé, Yann Batiste Pequignot 等AAAI 2025 · 被引用 6 次
- Frozen Language Models Are Gradient Coherence Rectifiers in Vision TransformersLichen Bai, Zixuan Xiong, Hai Lin, Guangwei Xu 等AAAI 2025 · 被引用 4 次
- DeepKD: A Deeply Decoupled and Denoised Knowledge Distillation TrainerHaiduo Huang, Jiangcheng Song, Yadong Zhang, Pengju RenNeurIPS 2025 · 被引用 3 次
- One-Step Generalization Ratio Guided Optimization for Domain GeneralizationSumin Cho, Dongwon Kim, Kwangsu KimICML 2025
- FD-MAGRPO: Functionality-Driven Multi-Agent Group Relative Policy Optimization for Analog-LDO SizingHaoning Jiang, Han Wu, Zhuoli Ouyang, Ziheng Wang 等AAAI 2026
它引用的顶会 Paper27
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh 等ICCV 2019 · 被引用 5,843 次
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang 等ICCV 2019 · 被引用 2,239 次
- Domain Generalization with MixStyleKaiyang Zhou, Yongxin Yang, Yu Qiao, Tao XiangICLR 2021 · 被引用 986 次
- Semi-Supervised Domain Adaptation via Minimax EntropyKuniaki Saito, Donghyun Kim, Stan Sclaroff, Trevor Darrell 等ICCV 2019 · 被引用 725 次
- Domain Generalization Using a Mixture of Multiple Latent DomainsToshihiko Matsuura, Tatsuya HaradaAAAI 2020 · 被引用 355 次
相关 Paper
- Understanding Why Neural Networks Generalize Well Through GSNR of ParametersJinlong Liu, Yunzhi Bai, Guoqing Jiang, Ting Chen 等ICLR 2020 · 被引用 60 次
- Learning Meta Face Recognition in Unseen DomainsJianzhu Guo, Xiangyu Zhu, Chenxu Zhao, Dong Cao 等CVPR 2020
- Regularized Fine-Grained Meta Face Anti-SpoofingRui Shao, Xiangyuan Lan, Pong C. YuenAAAI 2020 · 被引用 185 次
- Generalizable Representation Learning for Mixture Domain Face Anti-SpoofingZhihong Chen, Taiping Yao, Kekai Sheng, Shouhong Ding 等AAAI 2021 · 被引用 116 次
- Meta Dropout: Learning to Perturb Latent Features for GeneralizationHaebeom Lee, Taewook Nam, Eunho Yang, Sung Ju HwangICLR 2020 · 被引用 59 次
