Reproducibility in Multiple Instance Learning: A Case For Algorithmic Unit Tests
Edward Raff, James Holt
摘要
Multiple Instance Learning (MIL) is a sub-domain of classification problems with positive and negative labels and a "bag" of inputs, where the label is positive if and only if a positive element is contained within the bag, and otherwise is negative. Training in this context requires associating the bag-wide label to instance-level information, and implicitly contains a causal assumption and asymmetry to the task (i.e., you can't swap the labels without changing the semantics). MIL problems occur in healthcare (one malignant cell indicates cancer), cyber security (one malicious executable makes an infected computer), and many other tasks. In this work, we examine five of the most prominent deep-MIL models and find that none of them respects the standard MIL assumption. They are able to learn anticorrelated instances, i.e., defaulting to "positive" labels until seeing a negative counter-example, which should not be possible for a correct MIL model. We suspect that enhancements and other works derived from these models will share the same issue. In any context in which these models are being used, this creates the potential for learning incorrect models, which creates risk of operational failure. We identify and demonstrate this problem via a proposed "algorithmic unit test", where we create synthetic datasets that can be solved by a MIL respecting model, and which clearly reveal learning that violates MIL assumptions. The five evaluated methods each fail one or more of these tests. This provides a model-agnostic way to identify violations of modeling assumptions, which we hope will be useful for future development and evaluation of MIL models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Are Multiple Instance Learning Algorithms Learnable for Instances?Jaeseok Jang, Hyuk-Yoon KwonNeurIPS 2024 · 被引用 13 次
- Cracking Instance Jigsaw Puzzles: An Alternative to Multiple Instance Learning for Whole Slide Image AnalysisXiwen Chen, Peijie Qiu, Wenhui Zhu, Hao Wang 等ICCV 2025 · 被引用 2 次
它引用的顶会 Paper9
- TransMIL: Transformer based Correlated Multiple Instance Learning for Whole Slide Image ClassificationZhuchen Shao, Hao Bian, Yang Chen, Yifeng Wang 等NeurIPS 2021 · 被引用 1,163 次
- Modern Hopfield Networks and Attention for Immune Repertoire ClassificationMichael Widrich, Bernhard Schäfl, Milena Pavlovic, Hubert Ramsauer 等NeurIPS 2020 · 被引用 152 次
- Loss-Based Attention for Deep Multiple Instance LearningXiaoshuang Shi, Fuyong Xing, Yuanpu Xie, Zizhao Zhang 等AAAI 2020 · 被引用 123 次
- Hyperparameter Optimization Is Deceiving Us, and How to Stop ItA. Feder Cooper, Yucheng Lu, Jessica Zosa Forde, Christopher De SaNeurIPS 2021 · 被引用 40 次
- OutfitNet: Fashion Outfit Recommendation with Attention-Based Multiple Instance LearningYusan Lin, Maryam Moosaei, Hao YangWWW 2020 · 被引用 36 次
相关 Paper
- Interventional Multi-Instance Learning with Deconfounded Instance-Level PredictionTiancheng Lin, Hongteng Xu, Canqian Yang, Yi XuAAAI 2022 · 被引用 34 次
- Predicting Lymph Node Metastasis Using Histopathological Images Based on Multiple Instance Learning With Deep Graph ConvolutionYu Zhao, Fan Yang, Yuqi Fang, Hailing Liu 等CVPR 2020
- Model Agnostic Interpretability for Multiple Instance LearningJoseph Early, Christine Evers, Sarvapali D. RamchurnICLR 2022 · 被引用 15 次
- Multiple-Instance Learning from Similar and Dissimilar BagsLei Feng, Senlin Shu, Yuzhou Cao, Lue Tao 等KDD 2021 · 被引用 11 次
- Every Error has Its Magnitude: Asymmetric Mistake Severity Training for Multiclass Multiple Instance LearningSungrae Hong, Jiwon Jeong, Jisu Shin, Donghee Han 等CVPR 2026
