Visualizing Deep Networks by Optimizing with Integrated Gradients
Zhongang Qi, Saeed Khorram, Fuxin Li
Abstract
Understanding and interpreting the decisions made by deep learning models is valuable in many domains. In computer vision, computing heatmaps from a deep network is a popular approach for visualizing and understanding deep networks. However, heatmaps that do not correlate with the network may mislead human, hence the performance of heatmaps in providing a faithful explanation to the underlying deep network is crucial. In this paper, we propose I-GOS, which optimizes for a heatmap so that the classification scores on the masked image would maximally decrease. The main novelty of the approach is to compute descent directions based on the integrated gradients instead of the normal gradient, which avoids local optima and speeds up convergence. Compared with previous approaches, our method can flexibly compute heatmaps at any resolution for different user needs. Extensive experiments on several benchmark datasets show that the heatmaps produced by our approach are more correlated with the decision of the underlying deep network, in comparison with other state-of-the-art approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fa36f7ec-6109-4f57-a1fc-7ed486ee1df5Cited by top-tier papers24
- SCOUTER: Slot Attention-based Classifier for Explainable Image RecognitionLiangzhi Li, Bowen Wang, Manisha Verma, Yuta Nakashima et al.ICCV 2021 · 66 citations
- One Explanation is Not Enough: Structured Attention Graphs for Image ClassificationVivswan Shitole, Fuxin Li, Minsuk Kahng, Prasad Tadepalli et al.NeurIPS 2021 · 51 citations
- Finding Discriminative Filters for Specific Degradations in Blind Super-ResolutionLiangbin Xie, Xintao Wang, Chao Dong, Zhongang Qi et al.NeurIPS 2021 · 46 citations
- Fine-Grained Neural Network Explanation by Identifying Input Features with Predictive InformationYang Zhang, Ashkan Khakzar, Yawei Li, Azade Farshad et al.NeurIPS 2021 · 33 citations
- A Novel Visual Interpretability for Deep Neural Networks by Optimizing Activation Maps with PerturbationQing-Long Zhang, Lu Rao, Yubin YangAAAI 2021 · 26 citations
Related papers
- IDGI: A Framework to Eliminate Explanation Noise from Integrated GradientsRuo Yang, Binghui Wang, Mustafa BilgicCVPR 2023
- Integrated Decision Gradients: Compute Your Attributions Where the Model Makes Its DecisionChase Walker, Sumit Kumar Jha, Kenny Chen, Rickard EwetzAAAI 2024 · 25 citations
- Guided Integrated Gradients: An Adaptive Path Method for Removing NoiseAndrei Kapishnikov, Subhashini Venugopalan, Besim Avci, Ben Wedin et al.CVPR 2021
- Visual Explanations via Iterated Integrated AttributionsOren Barkan, Yehonatan Elisha, Yuval Asher, Amit Eshel et al.ICCV 2023 · 33 citations
- Manifold-Aligned Guided Integrated Gradients for Reliable Feature AttributionSoyeon Kim, Seongwoo Lim, Kyowoon Lee, Jaesik ChoiICML 2026 · 2 citations
