A Peek Into the Reasoning of Neural Networks: Interpreting With Structural Visual Concepts
Yunhao Ge, Yao Xiao, Zhi Xu, Meng Zheng, Srikrishna Karanam, Terrence Chen, Laurent Itti, Ziyan Wu
Abstract
Despite substantial progress in applying neural networks (NN) to a wide variety of areas, they still largely suffer from a lack of transparency and interpretability. While recent developments in explainable artificial intelligence attempt to bridge this gap (e.g., by visualizing the correlation between input pixels and final outputs), these approaches are limited to explaining low-level relationships, and crucially, do not provide insights on error correction. In this work, we propose a framework (VRX) to interpret classification NNs with intuitive structural visual concepts. Given a trained classification model, the proposed VRX extracts relevant class-specific visual concepts and organizes them using structural concept graphs (SCG) based on pairwise concept relationships. By means of knowledge distillation, we show VRX can take a step towards mimicking the reasoning process of NNs and provide logical, concept-level explanations for final model decisions. With extensive experiments, we empirically show VRX can meaningfully answer "why" and "why not" questions about the prediction, providing easy-to-understand insights about the reasoning process. We also show that these insights can potentially provide guidance on improving NN's performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cab549aa-481a-497a-88cb-ae9dedd94097Cited by top-tier papers11
- ConceptExplainer: Interactive Explanation for Deep Neural Networks from a Concept PerspectiveJinbin Huang, Aditi Mishra, Bum Chul Kwon, Chris BryanIEEE VIS 2022 · 46 citations
- Explaining Deep Convolutional Neural Networks via Latent Visual-Semantic Filter AttentionYu Yang, Seungbae Kim, Jungseock JooCVPR 2022 · 11 citations
- Knowledge-Aware Neuron Interpretation for Scene ClassificationYong Guan, Freddy Lécué, Jiaoyan Chen, Ru Li et al.AAAI 2024 · 3 citations
- Enhance Sketch Recognition's Explainability via Semantic Component-Level ParsingGuangming Zhu, Siyuan Wang, Tianci Wu, Liang ZhangAAAI 2024 · 2 citations
- Schema Inference for Interpretable Image ClassificationHaofei Zhang, Mengqi Xue, Xiaokang Liu, Kaixuan Chen et al.ICLR 2023 · 1 citation
Builds on1
Related papers
- A Knowledge Distillation-Based Approach to Enhance Transparency of Classifier ModelsYuchen Jiang, Xinyuan Zhao, Yihang Wu, Ahmad ChaddadAAAI 2025 · 5 citations
- ConEx: Human-Interpretable Saliency Maps via Concept-Aware AttributionYehonatan Elisha, Oren Barkan, Ziv Haddad, Noam KoenigsteinICML 2026
- Entropy-Based Logic Explanations of Neural NetworksPietro Barbiero, Gabriele Ciravegna, Francesco Giannini, Pietro Lió et al.AAAI 2022 · 97 citations
- "Why Not Other Classes?": Towards Class-Contrastive Back-Propagation ExplanationsYipei Wang, Xiaoqian WangNeurIPS 2022 · 17 citations
- Concept-based Explanation for Fine-grained Images and Its Application in Infectious Keratitis ClassificationZhengqing Fang, Kun Kuang, Yuxiao Lin, Fei Wu et al.ACM MM 2020 · 25 citations
