Balancing Bias and Variance for Active Weakly Supervised Learning
Hitesh Sapkota, Qi Yu
Abstract
As a widely used weakly supervised learning scheme, modern multiple instance learning (MIL) models achieve competitive performance at the bag level. However, instance-level prediction, which is essential for many important applications, remains largely unsatisfactory. We propose to conduct novel active deep multiple instance learning that samples a small subset of informative instances for annotation, aiming to significantly boost the instance-level prediction. A variance regularized loss function is designed to properly balance the bias and variance of instance-level predictions, aiming to effectively accommodate the highly imbalanced instance distribution in MIL and other fundamental challenges. Instead of directly minimizing the variance regularized loss that is non-convex, we optimize a distributionally robust bag level likelihood as its convex surrogate. The robust bag likelihood provides a good approximation of the variance based MIL loss with a strong theoretical guarantee. It also automatically balances bias and variance, making it effective to identify the potentially positive instances to support active sampling. The robust bag likelihood can be naturally integrated with a deep architecture to support deep model training using mini-batches of positive-negative bag pairs. Finally, a novel P-F sampling function is developed that combines a probability vector and predicted instance scores, obtained by optimizing the robust bag likelihood. By leveraging the key MIL assumption, the sampling function can explore the most challenging bags and effectively detect their positive instances for annotation, which significantly improves the instance-level prediction. Experiments conducted over multiple real-world datasets clearly demonstrate the state-of-the-art instance-level prediction achieved by the proposed model.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a533b52b-bcbf-4bed-9aad-d7da9e1fbd33Cited by top-tier papers1
Ask how each one uses itBuilds on3
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford et al.ICLR 2020 · 974 citations
- Reinforced active learning for image segmentationArantxa Casanova, Pedro O. Pinheiro, Negar Rostamzadeh, Christopher J. PalICLR 2020 · 127 citations
- Multiple Instance Active Learning for Object DetectionTianning Yuan, Fang Wan, Mengying Fu, Jianzhuang Liu et al.CVPR 2021
Related papers
- Loss-Based Attention for Deep Multiple Instance LearningXiaoshuang Shi, Fuyong Xing, Yuanpu Xie, Zizhao Zhang et al.AAAI 2020 · 123 citations
- Multiple-Instance Learning from Similar and Dissimilar BagsLei Feng, Senlin Shu, Yuzhou Cao, Lue Tao et al.KDD 2021 · 11 citations
- Are Multiple Instance Learning Algorithms Learnable for Instances?Jaeseok Jang, Hyuk-Yoon KwonNeurIPS 2024 · 13 citations
- Provable Multi-instance Deep AUC Maximization with Stochastic PoolingDixian Zhu, Bokun Wang, Zhi Chen, Yaxing Wang et al.ICML 2023 · 5 citations
- Predicting Lymph Node Metastasis Using Histopathological Images Based on Multiple Instance Learning With Deep Graph ConvolutionYu Zhao, Fan Yang, Yuqi Fang, Hailing Liu et al.CVPR 2020
