Explaining Neural Networks Semantically and Quantitatively
Runjin Chen, Hao Chen, Ge Huang, Jie Ren, Quanshi Zhang
Abstract
This paper 1 presents a method to explain the knowledge encoded in a convolutional neural network (CNN) quantitatively and semantically. The analysis of the specific rationale of each prediction made by the CNN presents a key issue of understanding neural networks, but it is also of significant practical values in certain applications. In this study, we propose to distill knowledge from the CNN into an explainable additive model, so that we can use the explainable model to provide a quantitative explanation for the CNN prediction. We analyze the typical bias-interpreting problem of the explainable model and develop prior losses to guide the learning of the explainable additive model. Experimental results have demonstrated the effectiveness of our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e93e5b00-2d35-40ec-931b-ec5ca842f128Cited by top-tier papers7
- Explainable Person Re-Identification with Attribute-guided Metric DistillationXiaodong Chen, Xinchen Liu, Wu Liu, Xiao-Ping Zhang et al.ICCV 2021 · 60 citations
- DISSECT: Disentangled Simultaneous Explanations via Concept TraversalsAsma Ghandeharioun, Been Kim, Chun-Liang Li, Brendan Jou et al.ICLR 2022 · 58 citations
- HyDRA: Hypergradient Data Relevance Analysis for Interpreting Deep Neural NetworksYuanyuan Chen, Boyang Li, Han Yu, Pengcheng Wu et al.AAAI 2021 · 50 citations
- Towards Automating Model Explanations with Certified Robustness GuaranteesMengdi Huai, Jinduo Liu, Chenglin Miao, Liuyi Yao et al.AAAI 2022 · 16 citations
- A Large-Scale Empirical Study on Improving the Fairness of Image Classification ModelsJunjie Yang, Jiajun Jiang, Zeyu Sun, Junjie ChenISSTA 2024 · 4 citations
Related papers
- A Peek Into the Reasoning of Neural Networks: Interpreting With Structural Visual ConceptsYunhao Ge, Yao Xiao, Zhi Xu, Meng Zheng et al.CVPR 2021
- A Knowledge Distillation-Based Approach to Enhance Transparency of Classifier ModelsYuchen Jiang, Xinyuan Zhao, Yihang Wu, Ahmad ChaddadAAAI 2025 · 5 citations
- Explaining Knowledge Distillation by Quantifying the KnowledgeXu Cheng, Zhefan Rao, Yilan Chen, Quanshi ZhangCVPR 2020
- Transferable Perturbations of Deep Feature DistributionsNathan Inkawhich, Kevin J. Liang, Lawrence Carin, Yiran ChenICLR 2020 · 100 citations
- Class Attention Transfer Based Knowledge DistillationZiyao Guo, Haonan Yan, Hui Li, Xiaodong LinCVPR 2023
