SelecMix: Debiased Learning by Contradicting-pair Sampling
Inwoo Hwang, Sangjun Lee, Yunhyeok Kwak, Seong Joon Oh, Damien Teney, Jin-Hwa Kim, Byoung-Tak Zhang
Abstract
Neural networks trained with ERM (empirical risk minimization) sometimes learn unintended decision rules, in particular when their training data is biased, i.e., when training labels are strongly correlated with undesirable features. To prevent a network from learning such features, recent methods augment training data such that examples displaying spurious correlations (i.e., bias-aligned examples) become a minority, whereas the other, bias-conflicting examples become prevalent. However, these approaches are sometimes difficult to train and scale to real-world data because they rely on generative models or disentangled representations. We propose an alternative based on mixup, a popular augmentation that creates convex combinations of training examples. Our method, coined SelecMix, applies mixup to contradicting pairs of examples, defined as showing either (i) the same label but dissimilar biased features, or (ii) different labels but similar biased features. Identifying such pairs requires comparing examples with respect to unknown biased features. For this, we utilize an auxiliary contrastive model with the popular heuristic that biased features are learned preferentially during training. Experiments on standard benchmarks demonstrate the effectiveness of the method, in particular when label noise complicates the identification of bias-conflicting examples.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e25cc0a6-b92d-4a0c-9ad1-d9f6e37f5d58Cited by top-tier papers12
- Selective Mixup Helps with Distribution Shifts, But Not (Only) because of MixupDamien Teney, Jindong Wang, Ehsan AbbasnejadICML 2024 · 9 citations
- A Simple Remedy for Dataset Bias via Self-Influence: A Mislabeled Sample PerspectiveYeonsung Jung, Jaeyun Song, June Yong Yang, Jin-Hwa Kim et al.NeurIPS 2024 · 7 citations
- Partition-and-Debias: Agnostic Biases Mitigation via A Mixture of Biases-Specific ExpertsJiaxuan Li, Duc Minh Vo, Hideki NakayamaICCV 2023 · 6 citations
- Navigate Beyond Shortcuts: Debiased Learning through the Lens of Neural CollapseYining Wang, Junjie Sun, Chenyue Wang, Mi Zhang et al.CVPR 2024 · 6 citations
- Diffusing DeBias: Synthetic Bias Amplification for Model DebiasingMassimiliano Ciranni, Vito Paolo Pastore, Roberto Di Via, Enzo Tartaglione et al.NeurIPS 2025 · 4 citations
Builds on22
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna et al.NeurIPS 2020 · 7,049 citations
- CutMix: Regularization Strategy to Train Strong Classifiers With Localizable FeaturesSangdoo Yun, Dongyoon Han, Sanghyuk Chun, Seong Joon Oh et al.ICCV 2019 · 5,843 citations
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 1,578 citations
- Just Train Twice: Improving Group Robustness without Training Group InformationEvan Zheran Liu, Behzad Haghgoo, Annie S. Chen, Aditi Raghunathan et al.ICML 2021 · 683 citations
Related papers
- BiasAdv: Bias-Adversarial Augmentation for Model DebiasingJongin Lim, Youngdong Kim, Byungjai Kim, Chanho Ahn et al.CVPR 2023
- Learning Debiased Representation via Disentangled Feature AugmentationJungsoo Lee, Eungyeup Kim, Juyoung Lee, Jihyeon Lee et al.NeurIPS 2021 · 203 citations
- Counterexample Contrastive Learning for Spurious Correlation EliminationJinqiang Wang, Rui Hu, Chaoquan Jiang, Rui Hu et al.ACM MM 2022 · 3 citations
- Provably Learning Diverse Features in Multi-View Data with Midpoint MixupMuthu Chidambaram, Xiang Wang, Chenwei Wu, Rong GeICML 2023 · 13 citations
- Echoes: Unsupervised Debiasing via Pseudo-bias Labeling in an Echo ChamberRui Hu, Yahan Tu, Jitao SangACM MM 2023 · 1 citation
