Reproducibility in Multiple Instance Learning: A Case For Algorithmic Unit Tests
Edward Raff, James Holt
Abstract
Multiple Instance Learning (MIL) is a sub-domain of classification problems with positive and negative labels and a "bag" of inputs, where the label is positive if and only if a positive element is contained within the bag, and otherwise is negative. Training in this context requires associating the bag-wide label to instance-level information, and implicitly contains a causal assumption and asymmetry to the task (i.e., you can't swap the labels without changing the semantics). MIL problems occur in healthcare (one malignant cell indicates cancer), cyber security (one malicious executable makes an infected computer), and many other tasks. In this work, we examine five of the most prominent deep-MIL models and find that none of them respects the standard MIL assumption. They are able to learn anticorrelated instances, i.e., defaulting to "positive" labels until seeing a negative counter-example, which should not be possible for a correct MIL model. We suspect that enhancements and other works derived from these models will share the same issue. In any context in which these models are being used, this creates the potential for learning incorrect models, which creates risk of operational failure. We identify and demonstrate this problem via a proposed "algorithmic unit test", where we create synthetic datasets that can be solved by a MIL respecting model, and which clearly reveal learning that violates MIL assumptions. The five evaluated methods each fail one or more of these tests. This provides a model-agnostic way to identify violations of modeling assumptions, which we hope will be useful for future development and evaluation of MIL models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4a118520-0d09-4dd2-a4d9-0e3052ce6424Cited by top-tier papers2
- Are Multiple Instance Learning Algorithms Learnable for Instances?Jaeseok Jang, Hyuk-Yoon KwonNeurIPS 2024 · 13 citations
- Cracking Instance Jigsaw Puzzles: An Alternative to Multiple Instance Learning for Whole Slide Image AnalysisXiwen Chen, Peijie Qiu, Wenhui Zhu, Hao Wang et al.ICCV 2025 · 2 citations
Builds on9
- TransMIL: Transformer based Correlated Multiple Instance Learning for Whole Slide Image ClassificationZhuchen Shao, Hao Bian, Yang Chen, Yifeng Wang et al.NeurIPS 2021 · 1,163 citations
- Modern Hopfield Networks and Attention for Immune Repertoire ClassificationMichael Widrich, Bernhard Schäfl, Milena Pavlovic, Hubert Ramsauer et al.NeurIPS 2020 · 152 citations
- Loss-Based Attention for Deep Multiple Instance LearningXiaoshuang Shi, Fuyong Xing, Yuanpu Xie, Zizhao Zhang et al.AAAI 2020 · 123 citations
- Hyperparameter Optimization Is Deceiving Us, and How to Stop ItA. Feder Cooper, Yucheng Lu, Jessica Zosa Forde, Christopher De SaNeurIPS 2021 · 40 citations
- OutfitNet: Fashion Outfit Recommendation with Attention-Based Multiple Instance LearningYusan Lin, Maryam Moosaei, Hao YangWWW 2020 · 36 citations
Related papers
- Interventional Multi-Instance Learning with Deconfounded Instance-Level PredictionTiancheng Lin, Hongteng Xu, Canqian Yang, Yi XuAAAI 2022 · 34 citations
- Predicting Lymph Node Metastasis Using Histopathological Images Based on Multiple Instance Learning With Deep Graph ConvolutionYu Zhao, Fan Yang, Yuqi Fang, Hailing Liu et al.CVPR 2020
- Model Agnostic Interpretability for Multiple Instance LearningJoseph Early, Christine Evers, Sarvapali D. RamchurnICLR 2022 · 15 citations
- Multiple-Instance Learning from Similar and Dissimilar BagsLei Feng, Senlin Shu, Yuzhou Cao, Lue Tao et al.KDD 2021 · 11 citations
- Every Error has Its Magnitude: Asymmetric Mistake Severity Training for Multiclass Multiple Instance LearningSungrae Hong, Jiwon Jeong, Jisu Shin, Donghee Han et al.CVPR 2026
