Dual Decomposition of Convex Optimization Layers for Consistent Attention in Medical Images
Tom Ron, Michal Weiler-Sagie, Tamir Hazan
Abstract
A key concern in integrating machine learning models in medicine is the ability to interpret their reasoning. Popular explainability methods have demonstrated satisfactory results in natural image recognition, yet in medical image analysis, many of these approaches provide partial and noisy explanations. Recently, attention mechanisms have shown compelling results both in their predictive performance and in their interpretable qualities. A fundamental trait of attention is that it leverages salient parts of the input which contribute to the model's prediction. To this end, our work focuses on the explanatory value of attention weight distributions. We propose a multi-layer attention mechanism that enforces consistent interpretations between attended convolutional layers using convex optimization. We apply duality to decompose the consistency constraints between the layers by reparameterizing their attention probability distributions. We further suggest learning the dual witness by optimizing with respect to our objective; thus, our implementation uses standard back-propagation, hence it is highly efficient. While preserving predictive performance, our proposed method leverages weakly annotated medical imaging data and provides complete and faithful explanations to the model's prediction.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8318c219-ece5-47ec-bd95-b235427f4a90Builds on10
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel et al.ICLR 2020 · 7,418 citations
- Attention Augmented Convolutional NetworksIrwan Bello, Barret Zoph, Quoc Le, Ashish Vaswani et al.ICCV 2019 · 1,149 citations
- Local Relation Networks for Image RecognitionHan Hu, Zheng Zhang, Zhenda Xie, Stephen LinICCV 2019 · 555 citations
- Understanding Deep Networks via Extremal Perturbations and Smooth MasksRuth Fong, Mandela Patrick, Andrea VedaldiICCV 2019 · 480 citations
Related papers
- Bayes-MIL: A New Probabilistic Perspective on Attention-based Multiple Instance Learning for Whole Slide ImagesYufei Cui, Ziquan Liu, Xiangyu Liu, Xue Liu et al.ICLR 2023
- A Knowledge Distillation-Based Approach to Enhance Transparency of Classifier ModelsYuchen Jiang, Xinyuan Zhao, Yihang Wu, Ahmad ChaddadAAAI 2025 · 5 citations
- MEDebiaser: A Human-AI Feedback System for Mitigating Bias in Multi-label Medical Image ClassificationShaohan Shi, Yuheng Shao, Haoran Jiang, Yunjie Yao et al.UIST 2025
- Deep Co-Attention Network for Multi-View Subspace LearningLecheng Zheng, Yu Cheng, Hongxia Yang, Nan Cao et al.WWW 2021 · 37 citations
- Additive MIL: Intrinsically Interpretable Multiple Instance Learning for PathologySyed Ashar Javed, Dinkar Juyal, Harshith Padigela, Amaro Taylor-Weiner et al.NeurIPS 2022 · 124 citations
