Empowering CAM-Based Methods with Capability to Generate Fine-Grained and High-Faithfulness Explanations
Changqing Qiu, Fusheng Jin, Yining Zhang
Abstract
Recently, the explanation of neural network models has garnered considerable research attention. In computer vision, CAM (Class Activation Map)-based methods and LRP (Layer-wise Relevance Propagation) method are two common explanation methods. However, since most CAM-based methods can only generate global weights, they can only generate coarse-grained explanations at a deep layer. LRP and its variants, on the other hand, can generate fine-grained explanations. But the faithfulness of the explanations is too low. To address these challenges, in this paper, we propose FG-CAM (Fine-Grained CAM), which extends CAM-based methods to enable generating fine-grained and high-faithfulness explanations. FG-CAM uses the relationship between two adjacent layers of feature maps with resolution differences to gradually increase the explanation resolution, while finding the contributing pixels and filtering out the pixels that do not contribute. Our method not only solves the shortcoming of CAM-based methods without changing their characteristics, but also generates fine-grained explanations that have higher faithfulness than LRP and its variants. We also present FG-CAM with denoising, which is a variant of FG-CAM and is able to generate less noisy explanations with almost no change in explanation faithfulness. Experimental results show that the performance of FG-CAM is almost unaffected by the explanation resolution. FG-CAM outperforms existing CAM-based methods significantly in both shallow and intermediate layers, and outperforms LRP and its variants significantly in the input layer. Our code is available at https://github.com/dongmo-qcq/FG-CAM .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 74e96e29-a2fd-4dc4-a0d2-1e3806542338Cited by top-tier papers1
Ask how each one uses itBuilds on5
- TS-CAM: Token Semantic Coupled Attention Map for Weakly Supervised Object LocalizationWei Gao, Fang Wan, Xingjia Pan, Zhiliang Peng et al.ICCV 2021 · 260 citations
- Invertible Concept-based Explanations for CNN Models with Non-negative Concept Activation VectorsRuihan Zhang, Prashan Madumal, Tim Miller, Krista A. Ehinger et al.AAAI 2021 · 140 citations
- Towards Better Explanations of Class Activation MappingHyungsik Jung, Youngrock OhICCV 2021 · 109 citations
- U-CAM: Visual Explanation Using Uncertainty Based Class Activation MapsBadri N. Patro, Mayank Lunayach, Shivansh Patel, Vinay P. NamboodiriICCV 2019 · 82 citations
- LFI-CAM: Learning Feature Importance for Better Visual ExplanationKwang Hee Lee, Chaewon Park, Junghyun Oh, Nojun KwakICCV 2021 · 38 citations
Related papers
- Relevance-CAM: Your Model Already Knows Where To LookJeong Ryong Lee, Sewon Kim, Inyong Park, Taejoon Eo et al.CVPR 2021
- Finer-CAM: Spotting the Difference Reveals Finer Details for Visual ExplanationZiheng Zhang, Jianyang Gu, Arpita Chowdhury, Zheda Mai et al.CVPR 2025
- Holistic-CAM: Ultra-lucid and Sanity Preserving Visual Interpretation in Holistic Stage of CNNsPengxu Chen, Huazhong Liu, Jihong Ding, Jiawen Luo et al.ACM MM 2024 · 8 citations
- Explaining Local, Global, And Higher-Order Interactions In Deep LearningSamuel Lerman, Charles Venuto, Henry A. Kautz, Chenliang XuICCV 2021 · 13 citations
- Explaining Convolutional Neural Networks through Attribution-Based Input Sampling and Block-Wise Feature AggregationSam Sattarzadeh, Mahesh Sudhakar, Anthony Lem, Shervin Mehryar et al.AAAI 2021 · 36 citations
