Interpreting Deep Neural Networks with Relative Sectional Propagation by Analyzing Comparative Gradients and Hostile Activations
Woo-Jeoung Nam, Jaesik Choi, Seong-Whan Lee
Abstract
The clear transparency of Deep Neural Networks (DNNs) is hampered by complex internal structures and nonlinear transformations along deep hierarchies. In this paper, we propose a new attribution method, Relative Sectional Propagation (RSP), for fully decomposing the output predictions with the characteristics of class-discriminative attributions and clear objectness. We carefully revisit some shortcomings of backpropagation-based attribution methods, which are trade-off relations in decomposing DNNs. We define hostile factor as an element that interferes with finding the attributions of the target and propagate it in a distinguishable way to overcome the non-suppressed nature of activated neurons. As a result, it is possible to assign the bi-polar relevance scores of the target (positive) and hostile (negative) attributions while maintaining each attribution aligned with the importance. We also present the purging techniques to prevent the decrement of the gap between the relevance scores of the target and hostile attributions during backward propagation by eliminating the conflicting units to channel attribution map. Therefore, our method makes it possible to decompose the predictions of DNNs with clearer class-discriminativeness and detailed elucidations of activation neurons compared to the conventional attribution methods. In a verified experimental environment, we report the results of the assessments: (i) Pointing Game, (ii) mIoU, and (iii) Model Sensitivity with PAS-CAL VOC 2007, MS COCO 2014, and ImageNet datasets. The results demonstrate that our method outperforms existing backward decomposition methods, including distinctive and intuitive visualizations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e3580f67-eae9-4047-ab04-c4e45f513624Cited by top-tier papers5
- Logic Rule Guided Attribution with Dynamic AblationJianqiao An, Yuandu Lai, Yahong HanAAAI 2022 · 4 citations
- Saliency strikes back: How filtering out high frequencies improves white-box explanationsSabine Muzellec, Thomas Fel, Victor Boutin, Léo Andéol et al.ICML 2024 · 4 citations
- Towards Better Visualizing the Decision Basis of Networks via Unfold and Conquer Attribution GuidanceJung-Ho Hong, Woo-Jeoung Nam, Kyu-Sung Jeon, Seong-Whan LeeAAAI 2023 · 3 citations
- Towards Fine-Grained Interpretability: Counterfactual Explanations for Misclassification with Saliency PartitionLintong Zhang, Kang Yin, Seong-Whan LeeCVPR 2025
- PINet: Improving the Stability of Prototype Networks via Phantasia-Inspired Uncertain RepresentationsHo Kyung Shin, Soeun Bae, Sang Min Kim, Byoung Chul Ko et al.AAAI 2026
Builds on3
- Understanding Deep Networks via Extremal Perturbations and Smooth MasksRuth Fong, Mandela Patrick, Andrea VedaldiICCV 2019 · 480 citations
- Relative Attributing Propagation: Interpreting the Comparative Contributions of Individual Units in Deep Neural NetworksWoo-Jeoung Nam, Shir Gur, Jaesik Choi, Lior Wolf et al.AAAI 2020 · 109 citations
- There and Back Again: Revisiting Backpropagation Saliency MethodsSylvestre-Alvise Rebuffi, Ruth Fong, Xu Ji, Andrea VedaldiCVPR 2020
Related papers
- Generating Attribution Maps with Disentangled Masked BackpropagationAdria Ruiz, Antonio Agudo, Francesc Moreno-NoguerICCV 2021 · 3 citations
- Mutual Information Preserving Back-propagation: Learn to Invert for Faithful AttributionHuiqi Deng, Na Zou, Weifu Chen, Guocan Feng et al.KDD 2021 · 3 citations
- Labeling Neural Representations with Inverse RecognitionKirill Bykov, Laura Kopf, Shinichi Nakajima, Marius Kloft et al.NeurIPS 2023 · 36 citations
- Distilled Gradient Aggregation: Purify Features for Input Attribution in the Deep Neural NetworkGiyoung Jeon, Haedong Jeong, Jaesik ChoiNeurIPS 2022 · 11 citations
- When Explanations Lie: Why Many Modified BP Attributions FailLeon Sixt, Maximilian Granz, Tim LandgrafICML 2020 · 147 citations
