DiCaP: Distribution-Calibrated Pseudo-labeling for Semi-Supervised Multi-Label Learning
Bo Han, Zhuoming Li, Xiaoyu Wang, Yaxin Hou, Hui Liu, Junhui Hou, Yuheng Jia
摘要
Semi-supervised multi-label learning (SSMLL) aims to address the challenge of limited labeled data in multi-label learning (MLL) by leveraging unlabeled data to improve the model’s performance. While pseudo-labeling has become a dominant strategy in SSMLL, most existing methods assign equal weights to all pseudo-labels regardless of their quality, which can amplify the impact of noisy or uncertain predictions and degrade the overall performance. In this paper, we theoretically verify that the optimal weight for a pseudo-label should reflect its correctness likelihood. Empirically, we observe that on the same dataset, the correctness likelihood distribution of unlabeled data remains stable, even as the number of labeled training samples varies. Building on this insight, we propose Distribution-Calibrated Pseudo-labeling (DiCaP), a correctness-aware framework that estimates posterior precision to calibrate pseudo-label weights. We further introduce a dual-thresholding mechanism to separate confident and ambiguous regions: confident samples are pseudo-labeled and weighted accordingly, while ambiguous ones are explored by unsupervised contrastive learning. Experiments conducted on multiple benchmark datasets verify that our method achieves consistent improvements, surpassing state-of-the-art methods by up to 4.27%.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper15
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 被引用 4,453 次
- Asymmetric Loss For Multi-Label ClassificationTal Ridnik, Emanuel Ben Baruch, Nadav Zamir, Asaf Noy 等ICCV 2021 · 被引用 778 次
- Class-Distribution-Aware Pseudo-Labeling for Semi-Supervised Multi-Label LearningMing-Kun Xie, Jiahao Xiao, Hao-Zhe Liu, Gang Niu 等NeurIPS 2023 · 被引用 55 次
- Dual Relation Semi-Supervised Multi-Label LearningLichen Wang, Yunyu Liu, Can Qin, Gan Sun 等AAAI 2020 · 被引用 47 次
相关 Paper
- Correlation-Induced Label Prior for Semi-Supervised Multi-Label LearningBiao Liu, Ning Xu, Xiangyu Fang, Xin GengICML 2024
- Asymmetric Beta Loss for Evidence-Based Safe Semi-Supervised Multi-Label LearningHao-Zhe Liu, Ming-Kun Xie, Chen-Chen Zong, Sheng-Jun HuangKDD 2024 · 被引用 1 次
- DAW: Exploring the Better Weighting Function for Semi-supervised Semantic SegmentationRui Sun, Huayu Mai, Tianzhu Zhang, Feng WuNeurIPS 2023 · 被引用 40 次
- In Defense of Pseudo-Labeling: An Uncertainty-Aware Pseudo-label Selection Framework for Semi-Supervised LearningMamshad Nayeem Rizve, Kevin Duarte, Yogesh S. Rawat, Mubarak ShahICLR 2021 · 被引用 630 次
- Rethinking Confidence Scores and Thresholds in Pseudolabeling-based SSLHarit Vishwakarma, Yi Chen, Satya Sai Srinath Namburi GNVV, Sui Jiet Tay 等ICML 2025
