Learning from positive and unlabeled examples -Finite size sample bounds
Farnam Mansouri, Shai Ben-David
Abstract
PU (Positive Unlabeled) learning is a variant of supervised classification learning in which the only labels revealed to the learner are of positively labeled instances. PU learning arises in many real-world applications. Most existing work relies on the simplifying assumptions that the positively labeled training data is drawn from the restriction of the data generating distribution to positively labeled instances and/or that the proportion of positively labeled points (a.k.a. the class prior) is known apriori to the learner. This paper provides a theoretical analysis of the statistical complexity of PU learning under a wider range of setups. Unlike most prior work, our study does not assume that the class prior is known to the learner. We prove upper and lower bounds on the required sample sizes (of both the positively labeled and the unlabeled samples).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 99b17cf2-058a-42e2-9646-db680c783132Cited by top-tier papers1
Ask how each one uses itBuilds on4
- Non-Negative Bregman Divergence Minimization for Deep Direct Density Ratio EstimationMasahiro Kato, Takeshi TeshimaICML 2021 · 53 citations
- Learning from Positive and Unlabeled Data with Arbitrary Positive ShiftZayd Hammoudeh, Daniel LowdNeurIPS 2020 · 53 citations
- GradPU: Positive-Unlabeled Learning via Gradient Penalty and Positive UpweightingSongmin Dai, Xiaoqiang Li, Yue Zhou, Xichen Ye et al.AAAI 2023 · 9 citations
- Smoothed Analysis of Learning from Positive SamplesJane H. Lee, Anay Mehrotra, Manolis ZampetakisSTOC 2026 · 2 citations
Related papers
- Recovering the Propensity Score from Biased Positive Unlabeled DataWalter Gerych, Thomas Hartvigsen, Luke Buquicchio, Emmanuel Agu et al.AAAI 2022 · 20 citations
- Class Prior Estimation with Biased Positives and Unlabeled ExamplesShantanu Jain, Justin Delano, Himanshu Sharma, Predrag RadivojacAAAI 2020 · 15 citations
- Rethinking Class-Prior Estimation for Positive-Unlabeled LearningYu Yao, Tongliang Liu, Bo Han, Mingming Gong et al.ICLR 2022 · 24 citations
- Balancing Positive and Negative Classification Error Rates in Positive-Unlabeled LearningXiming Li, Yuanchao Dai, Bing Wang, Changchun Li et al.NeurIPS 2025 · 3 citations
- Dist-PU: Positive-Unlabeled Learning from a Label Distribution PerspectiveYunrui Zhao, Qianqian Xu, Yangbangyan Jiang, Peisong Wen et al.CVPR 2022 · 47 citations
