PEER Pressure: Model-to-Model Regularization for Single Source Domain Generalization
Dong Kyu Cho, Inwoo Hwang, Sanghack Lee
摘要
Data augmentation is a popular tool for single source domain generalization, which expands the source domain by generating simulated ones, improving generalization on unseen target domains. In this work, we show that the performance of such augmentation-based methods in the target domains universally fluctuates during training, posing challenges in model selection under realistic scenarios. We argue that the fluctuation stems from the inability of the model to accumulate the knowledge learned from diverse augmentations, exacerbating feature distortion during training. Based on this observation, we propose a novel generalization method, coined Parameter-Space Ensemble with Entropy Regularization (PEER), that uses a proxy model to learn the augmented data on behalf of the main model. The main model is updated by averaging its parameters with the proxy model, progressively accumulating knowledge over the training steps. Maximizing the mutual information between the output representations of the two models guides the learning process of the proxy model, mitigating feature distortion during training. Experimental results demonstrate the effectiveness of PEER in reducing the OOD performance fluctuation and enhancing generalization across various datasets, including PACS, Digits, Office-Home, and VLCS. Notably, our method with simple random augmentation achieves state-of-the-art performance, surpassing prior approaches on sDG that utilize complex data augmentation strategies.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Forget Forgetting: Continual Learning in a World of Abundant MemoryDongkyu Cho, Taesup Moon, Rumi Chunara, Kyunghyun Cho 等ICLR 2026 · 被引用 9 次
- Modality-Balanced Collaborative Distillation for Multi-Modal Domain GeneralizationXiaohan Wang, Zhangtao Cheng, Ting Zhong, Leiting Chen 等AAAI 2026 · 被引用 3 次
- Expert-guided Clinical Text Augmentation via Query-Based Model CollaborationDongkyu Cho, Miao Zhang, Gregory Lyng, Rumi ChunaraICML 2026
- Spectral Property-Driven Data Augmentation for Hyperspectral Single-Source Domain GeneralizationTaiqin Chen, Yifeng Wang, Xiaochen Feng, Zhilin Zhu 等AAAI 2026
它引用的顶会 Paper26
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun 等ICML 2021 · 被引用 2,942 次
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs 等ICML 2022 · 被引用 1,464 次
- Fine-Tuning can Distort Pretrained Features and Underperform Out-of-DistributionAnanya Kumar, Aditi Raghunathan, Robbie Matthew Jones, Tengyu Ma 等ICLR 2022 · 被引用 911 次
- Linear Mode Connectivity and the Lottery Ticket HypothesisJonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, Michael CarbinICML 2020 · 被引用 750 次
相关 Paper
- A Simple Feature Augmentation for Domain GeneralizationPan Li, Da Li, Wei Li, Shaogang Gong 等ICCV 2021 · 被引用 242 次
- Practical Single Domain Generalization via Training-time and Test-time LearningShuai Yang, Zhen Zhang, Lichuan GuKDD 2024 · 被引用 3 次
- Adversarial Teacher-Student Representation Learning for Domain GeneralizationFu-En Yang, Yuan-Chia Cheng, Zu-Yun Shiau, Yu-Chiang Frank WangNeurIPS 2021 · 被引用 83 次
- PhysAug: A Physical-guided and Frequency-based Data Augmentation for Single-Domain Generalized Object DetectionXiaoran Xu, Jiangang Yang, Wenhui Shi, Siyuan Ding 等AAAI 2025 · 被引用 15 次
- Ensemble of Averages: Improving Model Selection and Boosting Performance in Domain GeneralizationDevansh Arpit, Huan Wang, Yingbo Zhou, Caiming XiongNeurIPS 2022 · 被引用 232 次
