Consistent Multi-Class Classification from Multiple Unlabeled Datasets
Zixi Wei, Senlin Shu, Yuzhou Cao, Hongxin Wei, Bo An, Lei Feng
摘要
Weakly supervised learning aims to construct effective predictive models from imperfectly labeled data. The recent trend of weakly supervised learning has focused on how to learn an accurate classifier from completely unlabeled data, given little supervised information such as class priors. In this paper, we consider a newly proposed weakly supervised learning problem called multi-class classification from multiple unlabeled datasets, where only multiple sets of unlabeled data and their class priors (i.e., the proportion of each class) are provided for training the classifier. To solve this problem, we first propose a classifier-consistent method (CCM) based on a probability transition function. However, CCM cannot guarantee risk consistency and lacks of purified supervision information during training. Therefore, we further propose a risk-consistent method (RCM) that progressively purifies supervision information during training by importance weighting. We provide comprehensive theoretical analyses for our methods to demonstrate the statistical consistency. Experimental results on multiple benchmark datasets across various settings demonstrate the superiority of our proposed methods.
• We provide comprehensive theoretical analyses for our proposed methods CCM and RCM to demonstrate their theoretical guarantees.
• We conduct extensive experiments on benchmark datasets with various settings. Experimental results demonstrate that CCM works well but RCM consistently outperforms CCM.
In this section, we introduce necessary notations, related studies, and the problem setting of our work.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper5
- Provably Consistent Partial-Label LearningLei Feng, Jiaqi Lv, Bo Han, Miao Xu 等NeurIPS 2020 · 被引用 188 次
- PiCO: Contrastive Label Disambiguation for Partial Label LearningHaobo Wang, Ruixuan Xiao, Yixuan Li, Lei Feng 等ICLR 2022 · 被引用 169 次
- Do We Need Zero Training Loss After Achieving Zero Training Error?Takashi Ishida, Ikko Yamane, Tomoya Sakai, Gang Niu 等ICML 2020 · 被引用 155 次
- Binary Classification from Multiple Unlabeled Datasets via Surrogate Set ClassificationNan Lu, Shida Lei, Gang Niu, Issei Sato 等ICML 2021 · 被引用 17 次
- Learning from Label Proportions: A Mutual Contamination FrameworkClayton Scott, Jianxin ZhangNeurIPS 2020 · 被引用 12 次
相关 Paper
- AUC Optimization from Multiple Unlabeled DatasetsZheng Xie, Yu Liu, Ming LiAAAI 2024 · 被引用 2 次
- Adversarial Multi Class Learning under Weak Supervision with Performance GuaranteesAlessio Mazzetto, Cyrus Cousins, Dylan Sam, Stephen H. Bach 等ICML 2021 · 被引用 39 次
- Learning with Complementary Labels Revisited: The Selected-Completely-at-Random Setting Is More PracticalWei Wang, Takashi Ishida, Yu-Jie Zhang, Gang Niu 等ICML 2024 · 被引用 12 次
- Can Class-Priors Help Single-Positive Multi-Label Learning?Biao Liu, Ning Xu, Jie Wang, Xin GengNeurIPS 2025
- A Universal Unbiased Method for Classification from Aggregate ObservationsZixi Wei, Lei Feng, Bo Han, Tongliang Liu 等ICML 2023 · 被引用 7 次
