Easy Learning from Label Proportions
Róbert Busa-Fekete, Heejin Choi, Travis Dick, Claudio Gentile, Andrés Muñoz Medina
摘要
We consider the problem of Learning from Label Proportions (LLP), a weakly supervised classification setup where instances are grouped into "bags", and only the frequency of class labels at each bag is available. Albeit, the objective of the learner is to achieve low task loss at an individual instance level. Here we propose EASYLLP: a flexible and simple-to-implement debiasing approach based on aggregate labels, which operates on arbitrary loss functions. Our technique allows us to accurately estimate the expected loss of an arbitrary model at an individual level. We showcase the flexibility of our approach by applying it to popular learning frameworks, like Empirical Risk Minimization (ERM) and Stochastic Gradient Descent (SGD) with provable guarantees on instance level performance. More concretely, we exhibit a variance reduction technique that makes the quality of LLP learning deteriorate only by a factor of k (k being bag size) in both ERM and SGD setups, as compared to full supervision. Finally, we validate our theoretical results on multiple datasets demonstrating our algorithm performs as well or better than previous LLP approaches in spite of its simplicity.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- PriorBoost: An Adaptive Algorithm for Learning from Aggregate ResponsesAdel Javanmard, Matthew Fahrbach, Vahab MirrokniICML 2024 · 被引用 6 次
- Learning from Label Proportions: Bootstrapping Supervised Learners via Belief PropagationShreyas Havaldar, Navodita Sharma, Shubhi Sareen, Karthikeyan Shanmugam 等ICLR 2024 · 被引用 5 次
- Learning from Aggregate responses: Instance Level versus Bag Level Loss FunctionsAdel Javanmard, Lin Chen, Vahab Mirrokni, Ashwinkumar Badanidiyuru 等ICLR 2024 · 被引用 3 次
- Auditing Privacy Mechanisms via Label Inference AttacksRóbert Busa-Fekete, Travis Dick, Claudio Gentile, Andrés Muñoz Medina 等NeurIPS 2024 · 被引用 3 次
- A Closer Look to Positive-Unlabeled Learning from Fine-grained Perspectives: An Empirical StudyYuanchao Dai, Zhengzhang Hou, Changchun Li, Yuanbo Xu 等NeurIPS 2025 · 被引用 2 次
它引用的顶会 Paper5
- Learning from Label Proportions by Learning with Label NoiseJianxin Zhang, Yutong Wang, Clayton ScottNeurIPS 2022 · 被引用 41 次
- Learnability of Linear Thresholds from Label ProportionsRishi SaketNeurIPS 2021 · 被引用 19 次
- Binary Classification from Multiple Unlabeled Datasets via Surrogate Set ClassificationNan Lu, Shida Lei, Gang Niu, Issei Sato 等ICML 2021 · 被引用 17 次
- Algorithms and Hardness for Learning Linear Thresholds from Label ProportionsRishi SaketNeurIPS 2022 · 被引用 15 次
- Learning from Label Proportions: A Mutual Contamination FrameworkClayton Scott, Jianxin ZhangNeurIPS 2020 · 被引用 12 次
相关 Paper
- Optimal Learning from Label Proportions with General Loss FunctionsLorne Applebaum, Travis Dick, Claudio Gentile, Haim Kaplan 等ICML 2026 · 被引用 1 次
- Nearly Optimal Sample Complexity for Learning with Label ProportionsRóbert Istvan Busa-Fekete, Travis Dick, Claudio Gentile, Haim Kaplan 等ICML 2025
- MixBag: Bag-Level Data Augmentation for Learning from Label ProportionsTakanori Asanomi, Shinnosuke Matsuo, Daiki Suehiro, Ryoma BiseICCV 2023 · 被引用 13 次
- Forming Auxiliary High-confident Instance-level Loss to Promote Learning from Label ProportionsTianhao Ma, Han Chen, Juncheng Hu, Yungang Zhu 等CVPR 2025
- Learning from Label Proportions via Proportional Value ClassificationTianhao Ma, Wei Wang, Ximing Li, Gang Niu 等ICLR 2026
