Learning Stable Classifiers by Transferring Unstable Features
Yujia Bao, Shiyu Chang, Regina Barzilay
摘要
While unbiased machine learning models are essential for many applications, bias is a human-defined concept that can vary across tasks. Given only input-label pairs, algorithms may lack sufficient information to distinguish stable (causal) features from unstable (spurious) features. However, related tasks often share similar biases -- an observation we may leverage to develop stable classifiers in the transfer setting. In this work, we explicitly inform the target classifier about unstable features in the source tasks. Specifically, we derive a representation that encodes the unstable features by contrasting different data environments in the source task. We achieve robustness by clustering data of the target task according to this representation and minimizing the worst-case risk across these clusters. We evaluate our method on both text and image classifications. Empirical results demonstrate that our algorithm is able to maintain robustness on the target task for both synthetically generated environments and real-world environments. Our code is available at https://github.com/YujiaBao/Tofu.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- SelecMix: Debiased Learning by Contradicting-pair SamplingInwoo Hwang, Sangjun Lee, Yunhyeok Kwak, Seong Joon Oh 等NeurIPS 2022 · 被引用 43 次
- A Whac-A-Mole Dilemma: Shortcuts Come in Multiples Where Mitigating One Amplifies OthersZhiheng Li, Ivan Evtimov, Albert Gordo, Caner Hazirbas 等CVPR 2023
它引用的顶会 Paper20
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 被引用 1,578 次
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang 等ICML 2021 · 被引用 1,163 次
- Environment Inference for Invariant LearningElliot Creager, Jörn-Henrik Jacobsen, Richard S. ZemelICML 2021 · 被引用 454 次
- Domain Generalization Using a Mixture of Multiple Latent DomainsToshihiko Matsuura, Tatsuya HaradaAAAI 2020 · 被引用 355 次
- No Subclass Left Behind: Fine-Grained Robustness in Coarse-Grained Classification ProblemsNimit Sharad Sohoni, Jared Dunnmon, Geoffrey Angus, Albert Gu 等NeurIPS 2020 · 被引用 316 次
相关 Paper
- Causal Transportability for Visual RecognitionChengzhi Mao, Kevin Xia, James Wang, Hao Wang 等CVPR 2022 · 被引用 27 次
- TACIT: A Target-Agnostic Feature Disentanglement Framework for Cross-Domain Text ClassificationRui Song, Fausto Giunchiglia, Yingji Li, Mingjie Tian 等AAAI 2024 · 被引用 10 次
- Examining and Combating Spurious Features under Distribution ShiftChunting Zhou, Xuezhe Ma, Paul Michel, Graham NeubigICML 2021 · 被引用 78 次
- Adv-SSL: Adversarial Self-Supervised Representation Learning with Theoretical GuaranteesChenguang Duan, Yuling Jiao, Huazhen Lin, Wensen Ma 等NeurIPS 2025 · 被引用 1 次
- Spuriosity Didn't Kill the Classifier: Using Invariant Predictions to Harness Spurious FeaturesCian Eastwood, Shashank Singh, Andrei Liviu Nicolicioiu, Marin Vlastelica Pogancic 等NeurIPS 2023 · 被引用 29 次
