A Novel Visual Interpretability for Deep Neural Networks by Optimizing Activation Maps with Perturbation
Qing-Long Zhang, Lu Rao, Yubin Yang
Abstract
Interpretability has been regarded as an essential component for deploying deep neural networks, in which the saliency-based method is one of the most prevailing interpretable approaches since it can generate individually intuitive heatmaps that highlight parts of the input image that are most important to the decision of the deep networks on a particular classification target. However, heatmaps generated by existing methods either contain little information to represent objects (perturbation-based methods) or cannot effectively locate multi-class objects (activation-based approaches). To address this issue, a two-stage framework for visualizing the interpretability of deep neural networks, called Activation Optimized with Perturbation (AOP), is designed to optimize activation maps generated by general activation-based methods with the help of perturbation-based methods. Finally, in order to obtain better explanations for different types of images, we further present an instance of the AOP framework, Smooth Integrated Gradient-based Class Activation Map (SIGCAM), which proposes a weighted GradCAM by applying the feature map as weight coefficients and employs I-GOS to optimize the base-mask generated by weighted GradCAM. Experimental results on common-used benchmarks, including deletion and insertion tests on ImageNet-1k, and pointing game tests on COCO2017, show that the proposed AOP and SIGCAM outperform the current state-of-the-art methods significantly by generating higher quality image-based saliency maps.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Interpretable3D: An Ad-Hoc Interpretable Classifier for 3D Point CloudsTuo Feng, Ruijie Quan, Xiaohan Wang, Wenguan Wang et al.AAAI 2024 · 31 citations
- "Is your explanation stable?": A Robustness Evaluation Framework for Feature AttributionYuyou Gan, Yuhao Mao, Xuhong Zhang, Shouling Ji et al.CCS 2022 · 13 citations
- Towards Better Visualizing the Decision Basis of Networks via Unfold and Conquer Attribution GuidanceJung-Ho Hong, Woo-Jeoung Nam, Kyu-Sung Jeon, Seong-Whan LeeAAAI 2023 · 3 citations
- FAM: Visual Explanations for the Feature Representations from Deep Convolutional NetworksYuxi Wu, Changhuai Chen, Jun Che, Shiliang PuCVPR 2022 · 1 citation
Builds on2
Related papers
- Explainable Models with Consistent InterpretationsVipin Pillai, Hamed PirsiavashAAAI 2021 · 46 citations
- Guided Integrated Gradients: An Adaptive Path Method for Removing NoiseAndrei Kapishnikov, Subhashini Venugopalan, Besim Avci, Ben Wedin et al.CVPR 2021
- Relevance-CAM: Your Model Already Knows Where To LookJeong Ryong Lee, Sewon Kim, Inyong Park, Taejoon Eo et al.CVPR 2021
- ODAM: Gradient-based Instance-Specific Visual Explanations for Object DetectionChenyang Zhao, Antoni B. ChanICLR 2023 · 4 citations
- Online Refinement of Low-level Feature Based Activation Map for Weakly Supervised Object LocalizationJinheng Xie, Cheng Luo, Xiangping Zhu, Ziqi Jin et al.ICCV 2021 · 61 citations
