InPL: Pseudo-labeling the Inliers First for Imbalanced Semi-supervised Learning
Zhuoran Yu, Yin Li, Yong Jae Lee
摘要
Recent state-of-the-art methods in imbalanced semi-supervised learning (SSL) rely on confidence-based pseudo-labeling with consistency regularization. To obtain high-quality pseudo-labels, a high confidence threshold is typically adopted. However, it has been shown that softmax-based confidence scores in deep networks can be arbitrarily high for samples far from the training data, and thus, the pseudo-labels for even high-confidence unlabeled samples may still be unreliable. In this work, we present a new perspective of pseudo-labeling for imbalanced SSL. Without relying on model confidence, we propose to measure whether an unlabeled sample is likely to be in-distribution''; i.e., close to the current training data. To decide whether an unlabeled sample is in-distribution'' or ``out-of-distribution'', we adopt the energy score from out-of-distribution detection literature. As training progresses and more unlabeled samples become in-distribution and contribute to training, the combined labeled and pseudo-labeled data can better approximate the true class distribution to improve the model. Experiments demonstrate that our energy-based pseudo-labeling method, InPL, albeit conceptually simple, significantly outperforms confidence-based methods on imbalanced SSL benchmarks. For example, it produces around 3% absolute accuracy improvement on CIFAR10-LT. When combined with state-of-the-art long-tailed SSL methods, further improvements are attained. In particular, in one of the most challenging scenarios, InPL achieves a 6.9% accuracy improvement over the best competitor.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Continuous Contrastive Learning for Long-Tailed Semi-Supervised RecognitionZi-Hao Zhou, Siyuan Fang, Zi-Jing Zhou, Tong Wei 等NeurIPS 2024 · 被引用 19 次
- Learnable Logit Adjustment for Imbalanced Semi-Supervised Learning Under Class Distribution MismatchHyuck Lee, Taemin Park, Heeyoung KimICCV 2025 · 被引用 1 次
- Learning Dynamics of Logits Debiasing for Long-Tailed Semi-Supervised LearningYue Cheng, Jiajun Zhang, Xiaohui Gao, Weiwei Xing 等ICLR 2026 · 被引用 1 次
- Sampling Control for Imbalanced Calibration in Semi-Supervised LearningSenmao Tian, Xiang Wei, Shunli ZhangAAAI 2026
- SeMi: When Imbalanced Semi-Supervised Learning Meets Mining Hard ExamplesYin Wang, Zixuan Wang, Hao Lu, Zhen Qin 等ACM MM 2025
它引用的顶会 Paper17
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang 等NeurIPS 2020 · 被引用 5,129 次
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 被引用 4,453 次
- Unsupervised Data Augmentation for Consistency TrainingQizhe Xie, Zihang Dai, Eduard H. Hovy, Thang Luong 等NeurIPS 2020 · 被引用 2,774 次
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 被引用 2,213 次
- FlexMatch: Boosting Semi-Supervised Learning with Curriculum Pseudo LabelingBowen Zhang, Yidong Wang, Wenxin Hou, Hao Wu 等NeurIPS 2021 · 被引用 1,389 次
相关 Paper
- Towards Realistic Long-Tailed Semi-Supervised Learning: Consistency is All You NeedTong Wei, Kai GanCVPR 2023
- CaliMatch: Adaptive Calibration for Improving Safe Semi-Supervised LearningJinsoo Bae, Seoung Bum Kim, Hyungrok DoICCV 2025 · 被引用 1 次
- Balanced Energy Regularization Loss for Out-of-distribution DetectionHyunjun Choi, Hawook Jeong, Jin Young ChoiCVPR 2023
- DASO: Distribution-Aware Semantics-Oriented Pseudo-label for Imbalanced Semi-Supervised LearningYoungtaek Oh, Dong-Jin Kim, In So KweonCVPR 2022 · 被引用 80 次
- Safe-Student for Safe Deep Semi-Supervised Learning with Unseen-Class Unlabeled DataRundong He, Zhongyi Han, Xiankai Lu, Yilong YinCVPR 2022 · 被引用 49 次
