Learning from Positive and Unlabeled Data without Explicit Estimation of Class Prior
Chenguang Zhang, Yuexian Hou, Yan Zhang
Abstract
Learning a classifier from positive and unlabeled data may occur in various applications. It differs from the standard classification problems by the absence of labeled negative examples in the training set. So far, two main strategies have typically been used for this issue: the likely negative examplesbased strategy and the class prior-based strategy, in which the likely negative examples or the class prior is required to be obtained in a preprocessing step. In this paper, a new strategy based on the Bhattacharyya coefficient is put forward, which formalizes this learning problem as an optimization problem and does not need a preprocessing step. We first show that with the given positive class conditional probability density function (PDF) and the mixture PDF of both the positive class and the negative class, the class prior can be estimated by minimizing the Bhattacharyya coefficient of the positive class with respect to the negative class. We then show how to use this result in an implicit mixture model of restricted Boltzmann machines to estimate the positive class conditional PDF and the negative class conditional PDF directly to obtain a classifier without the explicit estimation of the class prior. Many experiments on real and synthetic datasets illustrated the superiority of the proposed approach.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 65142624-bd04-4bab-bf1e-75815721dfc2Cited by top-tier papers4
- Predictive Adversarial Learning from Positive and Unlabeled DataWenpeng Hu, Ran Le, Bing Liu, Feng Ji et al.AAAI 2021 · 56 citations
- PULNS: Positive-Unlabeled Learning with Effective Negative Sample SelectorChuan Luo, Pu Zhao, Chen Chen, Bo Qiao et al.AAAI 2021 · 48 citations
- Positive Distribution Pollution: Rethinking Positive Unlabeled Learning from a Unified PerspectiveQianqiao Liang, Mengying Zhu, Yan Wang, Xiuyuan Wang et al.AAAI 2023 · 4 citations
- Regression with Sensor Data Containing Incomplete ObservationsTakayuki Katsuki, Takayuki OsogamiICML 2023 · 1 citation
Related papers
- A Variational Approach for Learning from Positive and Unlabeled DataHui Chen, Fangqing Liu, Yin Wang, Liyue Zhao et al.NeurIPS 2020 · 76 citations
- Rethinking Class-Prior Estimation for Positive-Unlabeled LearningYu Yao, Tongliang Liu, Bo Han, Mingming Gong et al.ICLR 2022 · 24 citations
- Fast Nonparametric Estimation of Class Proportions in the Positive-Unlabeled Classification SettingDaniel Zeiberg, Shantanu Jain, Predrag RadivojacAAAI 2020 · 23 citations
- Balancing Positive and Negative Classification Error Rates in Positive-Unlabeled LearningXiming Li, Yuanchao Dai, Bing Wang, Changchun Li et al.NeurIPS 2025 · 3 citations
- Class Prior Estimation with Biased Positives and Unlabeled ExamplesShantanu Jain, Justin Delano, Himanshu Sharma, Predrag RadivojacAAAI 2020 · 15 citations
