Robust Positive-Unlabeled Learning via Noise Negative Sample Self-correction
Zhangchi Zhu, Lu Wang, Pu Zhao, Chao Du, Wei Zhang, Hang Dong, Bo Qiao, Qingwei Lin, Saravan Rajmohan, Dongmei Zhang
Abstract
Learning from positive and unlabeled data is known as positive-unlabeled (PU) learning in literature and has attracted much attention in recent years. One common approach in PU learning is to sample a set of pseudo-negatives from the unlabeled data using ad-hoc thresholds so that conventional supervised methods can be applied with both positive and negative samples. Owing to the label uncertainty among the unlabeled data, errors of misclassifying unlabeled positive samples as negative samples inevitably appear and may even accumulate during the training processes. Those errors often lead to performance degradation and model instability. To mitigate the impact of label uncertainty and improve the robustness of learning with positive and unlabeled data, we propose a new robust PU learning method with a training strategy motivated by the nature of human learning: easy cases should be learned first. Similar intuition has been utilized in curriculum learning to only use easier cases in the early stage of training before introducing more complex cases. Specifically, we utilize a novel ''hardness'' measure to distinguish unlabeled samples with a high chance of being negative from unlabeled samples with large label noise. An iterative training strategy is then implemented to fine-tune the selection of negative samples during the training process in an iterative manner to include more ''easy'' samples in the early stage of training. Extensive experimental validations over a wide range of learning tasks show that this approach can effectively improve the accuracy and stability of learning with positive and unlabeled data. Our code is available at https://github.com/woriazzc/Robust-PU.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b973cd05-9e98-476b-895d-a9e0edc99c51Cited by top-tier papers6
- Positive and Unlabeled Learning with Controlled Probability Boundary FenceChangchun Li, Yuanchao Dai, Lei Feng, Ximing Li et al.ICML 2024 · 8 citations
- Balancing Positive and Negative Classification Error Rates in Positive-Unlabeled LearningXiming Li, Yuanchao Dai, Bing Wang, Changchun Li et al.NeurIPS 2025 · 3 citations
- A Closer Look to Positive-Unlabeled Learning from Fine-grained Perspectives: An Empirical StudyYuanchao Dai, Zhengzhang Hou, Changchun Li, Yuanbo Xu et al.NeurIPS 2025 · 2 citations
- Accessible, Realistic, and Fair Evaluation of Positive-Unlabeled Learning AlgorithmsWei Wang, Dong-Dong Wu, Ming Li, Jingxiong Zhang et al.ICLR 2026 · 2 citations
- PU-BENCH: A Unified Benchmark for Rigorous and Reproducible PU LearningQiuyi Chen, Haiyang Zhang, Leqi Zhang, Changchun Li et al.ICLR 2026
Builds on6
- Dynamic Curriculum Learning for Imbalanced Data ClassificationYiru Wang, Weihao Gan, Jie Yang, Wei Wu et al.ICCV 2019 · 263 citations
- Robust Curriculum Learning: from clean label detection to noisy label self-correctionTianyi Zhou, Shengjie Wang, Jeff A. BilmesICLR 2021 · 111 citations
- Self-PU: Self Boosted and Calibrated Positive-Unlabeled TrainingXuxi Chen, Wuyang Chen, Tianlong Chen, Ye Yuan et al.ICML 2020 · 100 citations
- PULNS: Positive-Unlabeled Learning with Effective Negative Sample SelectorChuan Luo, Pu Zhao, Chen Chen, Bo Qiao et al.AAAI 2021 · 48 citations
- Dist-PU: Positive-Unlabeled Learning from a Label Distribution PerspectiveYunrui Zhao, Qianqian Xu, Yangbangyan Jiang, Peisong Wen et al.CVPR 2022 · 47 citations
Related papers
- Split-PU: Hardness-aware Training Strategy for Positive-Unlabeled LearningChengming Xu, Chen Liu, Siqian Yang, Yabiao Wang et al.ACM MM 2022 · 4 citations
- Learning from Positive and Unlabeled Data with Arbitrary Positive ShiftZayd Hammoudeh, Daniel LowdNeurIPS 2020 · 53 citations
- Who Is Your Right Mixup Partner in Positive and Unlabeled LearningChangchun Li, Ximing Li, Lei Feng, Jihong OuyangICLR 2022 · 36 citations
- Predictive Adversarial Learning from Positive and Unlabeled DataWenpeng Hu, Ran Le, Bing Liu, Feng Ji et al.AAAI 2021 · 56 citations
- Recovering the Propensity Score from Biased Positive Unlabeled DataWalter Gerych, Thomas Hartvigsen, Luke Buquicchio, Emmanuel Agu et al.AAAI 2022 · 20 citations
