Distilled Gradient Aggregation: Purify Features for Input Attribution in the Deep Neural Network
Giyoung Jeon, Haedong Jeong, Jaesik Choi
Abstract
Measuring the attribution of input features toward the model output is one of the popular post-hoc explanations on the Deep Neural Networks (DNNs). Among various approaches to compute the attribution, the gradient-based methods are widely used to generate attributions, because of its ease of implementation and the model-agnostic characteristic. However, existing gradient integration methods such as Integrated Gradients (IG) suffer from (1) the noisy attributions which cause the unreliability of the explanation, and (2) the selection for the integration path which determines the quality of explanations. FullGrad (FG) is an another approach to construct the reliable attributions by focusing the locality of piece-wise linear network with the bias gradient. Although FG has shown reasonable performance for the given input, as the shortage of the global property, FG is vulnerable to the small perturbation, while IG which includes the exploration over the input space is robust. In this work, we design a new input attribution method which adopt the strengths of both local and global attributions. In particular, we propose a novel approach to distill input features using weak and extremely positive contributor masks. We aggregate the intermediate local attributions obtained from the distillation sequence to provide reliable attribution. We perform the quantitative evaluation compared to various attribution methods and show that our method outperforms others. We also provide the qualitative result that our method obtains object-aligned and sharp attribution heatmap. * Equal Contribution 36th Conference on Neural Information Processing Systems (NeurIPS 2022). (a) Distilled Gradient Aggregation (DGA) (b) Modules (detailed) Mask Extraction Local Attribution
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c8f328f5-8b55-414a-b014-365ee45f2636Cited by top-tier papers2
- MFABA: A More Faithful and Accelerated Boundary-Based Attribution Method for Deep Neural NetworksZhiyu Zhu, Huaming Chen, Jiayu Zhang, Xinyi Wang et al.AAAI 2024 · 16 citations
- Spectral Integrated Gradients for Coarse-to-Fine Feature AttributionSoyeon Kim, Seongwoo Lim, Kyowoon Lee, Jaesik ChoiKDD 2026 · 2 citations
Builds on5
- Understanding Deep Networks via Extremal Perturbations and Smooth MasksRuth Fong, Mandela Patrick, Andrea VedaldiICCV 2019 · 480 citations
- XRAI: Better Attributions Through RegionsAndrei Kapishnikov, Tolga Bolukbasi, Fernanda B. Viégas, Michael TerryICCV 2019 · 251 citations
- Visualizing Deep Networks by Optimizing with Integrated GradientsZhongang Qi, Saeed Khorram, Fuxin LiAAAI 2020 · 149 citations
- Relative Attributing Propagation: Interpreting the Comparative Contributions of Individual Units in Deep Neural NetworksWoo-Jeoung Nam, Shir Gur, Jaesik Choi, Lior Wolf et al.AAAI 2020 · 109 citations
- Guided Integrated Gradients: An Adaptive Path Method for Removing NoiseAndrei Kapishnikov, Subhashini Venugopalan, Besim Avci, Ben Wedin et al.CVPR 2021
Related papers
- Beyond Single Path Integrated Gradients for Reliable Input Attribution via Randomized Path SamplingGiyoung Jeon, Haedong Jeong, Jaesik ChoiICCV 2023 · 3 citations
- Manifold-Aligned Guided Integrated Gradients for Reliable Feature AttributionSoyeon Kim, Seongwoo Lim, Kyowoon Lee, Jaesik ChoiICML 2026 · 2 citations
- Local Path Integration for AttributionPeiyu Yang, Naveed Akhtar, Zeyi Wen, Ajmal MianAAAI 2023 · 16 citations
- Towards Better Understanding Attribution MethodsSukrut Rao, Moritz Böhle, Bernt SchieleCVPR 2022 · 32 citations
- Denoising Diffusion Path: Attribution Noise Reduction with An Auxiliary Diffusion ModelYiming Lei, Zilong Li, Junping Zhang, Hongming ShanNeurIPS 2024 · 9 citations
