Self-training Avoids Using Spurious Features Under Domain Shift
Yining Chen, Colin Wei, Ananya Kumar, Tengyu Ma
摘要
In unsupervised domain adaptation, existing theory focuses on situations where the source and target domains are close. In practice, conditional entropy minimization and pseudo-labeling work even when the domain shifts are much larger than those analyzed by existing theory. We identify and analyze one particular setting where the domain shift can be large, but these algorithms provably work: certain spurious features correlate with the label in the source domain but are independent of the label in the target. Our analysis considers linear classification where the spurious features are Gaussian and the non-spurious features are a mixture of log-concave distributions. For this setting, we prove that entropy minimization on unlabeled target data will avoid using the spurious feature if initialized with a decently accurate source classifier, even though the objective is non-convex and contains multiple bad local minima using the spurious features. We verify our theory for spurious domain shift tasks on semi-synthetic Celeb-A and MNIST datasets. Our results suggest that practitioners collect and self-train on large, diverse datasets to reduce biases in classifiers even if labeling is impractical.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper30
- Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak SupervisionCollin Burns, Pavel Izmailov, Jan Hendrik Kirchner, Bowen Baker 等ICML 2024 · 被引用 443 次
- Theoretical Analysis of Self-Training with Deep Networks on Unlabeled DataColin Wei, Kendrick Shen, Yining Chen, Tengyu MaICLR 2021 · 被引用 261 次
- Cycle Self-Training for Domain AdaptationHong Liu, Jianmin Wang, Mingsheng LongNeurIPS 2021 · 被引用 236 次
- Test Time Adaptation via Conjugate Pseudo-labelsSachin Goyal, Mingjie Sun, Aditi Raghunathan, J. Zico KolterNeurIPS 2022 · 被引用 152 次
- FreeMatch: Self-adaptive Thresholding for Semi-supervised LearningYidong Wang, Hao Chen, Qiang Heng, Wenxin Hou 等ICLR 2023 · 被引用 139 次
它引用的顶会 Paper6
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang 等NeurIPS 2020 · 被引用 5,129 次
- Confidence Regularized Self-TrainingYang Zou, Zhiding Yu, Xiaofeng Liu, B. V. K. Vijaya Kumar 等ICCV 2019 · 被引用 901 次
- ReMixMatch: Semi-Supervised Learning with Distribution Matching and Augmentation AnchoringDavid Berthelot, Nicholas Carlini, Ekin D. Cubuk, Alex Kurakin 等ICLR 2020 · 被引用 469 次
- An Investigation of Why Overparameterization Exacerbates Spurious CorrelationsShiori Sagawa, Aditi Raghunathan, Pang Wei Koh, Percy LiangICML 2020 · 被引用 436 次
- Understanding Self-Training for Gradual Domain AdaptationAnanya Kumar, Tengyu Ma, Percy LiangICML 2020 · 被引用 266 次
相关 Paper
- SENTRY: Selective Entropy Optimization via Committee Consistency for Unsupervised Domain AdaptationViraj Prabhu, Shivam Khare, Deeksha Kartik, Judy HoffmanICCV 2021 · 被引用 155 次
- CASUAL: Conditional Support Alignment for Domain Adaptation with Label ShiftAnh T. Nguyen, Lam Tran, Anh Tong, Tuan-Duy H. Nguyen 等AAAI 2025 · 被引用 3 次
- Connect, Not Collapse: Explaining Contrastive Learning for Unsupervised Domain AdaptationKendrick Shen, Robbie M. Jones, Ananya Kumar, Sang Michael Xie 等ICML 2022 · 被引用 102 次
- Implicit Class-Conditioned Domain Alignment for Unsupervised Domain AdaptationXiang Jiang, Qicheng Lao, Stan Matwin, Mohammad HavaeiICML 2020 · 被引用 129 次
- Semi-Supervised Domain Adaptation via Minimax EntropyKuniaki Saito, Donghyun Kim, Stan Sclaroff, Trevor Darrell 等ICCV 2019 · 被引用 725 次
