Knowledge-Aware Neuron Interpretation for Scene Classification
Yong Guan, Freddy Lécué, Jiaoyan Chen, Ru Li, Jeff Z. Pan
Abstract
Although neural models have achieved remarkable performance, they still encounter doubts due to the intransparency. To this end, model prediction explanation is attracting more and more attentions. However, current methods rarely incorporate external knowledge and still suffer from three limitations: (1) Neglecting concept completeness. Merely selecting concepts may not sufficient for prediction. (2) Lacking concept fusion. Failure to merge semantically-equivalent concepts. ( 3 ) Difficult in manipulating model behavior. Lack of verification for explanation on original model. To address these issues, we propose a novel knowledge-aware neuron interpretation framework to explain model predictions for image scene classification. Specifically, for concept completeness, we present core concepts of a scene based on knowledge graph, ConceptNet, to gauge the completeness of concepts. Our method, incorporating complete concepts, effectively provides better prediction explanations compared to baselines. Furthermore, for concept fusion, we introduce a knowledge graph-based method known as Concept Filtering, which produces over 23% point gain on neuron behaviors for neuron interpretation. At last, we propose Model Manipulation, which aims to study whether the core concepts based on ConceptNet could be employed to manipulate model behavior. The results show that core concepts can effectively improve the performance of original model by over 26%.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f2e6c708-0cc3-40a8-8038-a33f50d9f2dcBuilds on5
- On Completeness-aware Concept-Based Explanations in Deep Neural NetworksChih-Kuan Yeh, Been Kim, Sercan Ömer Arik, Chun-Liang Li et al.NeurIPS 2020 · 390 citations
- Compositional Explanations of NeuronsJesse Mu, Jacob AndreasNeurIPS 2020 · 229 citations
- Rethinking Interpretation: Input-Agnostic Saliency Mapping of Deep Visual ClassifiersNaveed Akhtar, Mohammad Amir Asim Khan JalwanaAAAI 2023 · 7 citations
- A Peek Into the Reasoning of Neural Networks: Interpreting With Structural Visual ConceptsYunhao Ge, Yao Xiao, Zhi Xu, Meng Zheng et al.CVPR 2021
- Alignment Rationale for Natural Language InferenceZhongtao Jiang, Yuanzhe Zhang, Zhao Yang, Jun Zhao et al.ACL 2021
Related papers
- eXpath: Explaining Knowledge Graph Link Prediction with Ontological Closed Path RulesYe Sun, Lei Shi, Yongxin TongVLDB 2025 · 3 citations
- Relational Concept Bottleneck ModelsPietro Barbiero, Francesco Giannini, Gabriele Ciravegna, Michelangelo Diligenti et al.NeurIPS 2024 · 21 citations
- Structure Your Data: Towards Semantic Graph CounterfactualsAngeliki Dimitriou, Maria Lymperaiou, Giorgos Filandrianos, Konstantinos Thomas et al.ICML 2024 · 7 citations
- Explaining Link Prediction Systems based on Knowledge Graph EmbeddingsAndrea Rossi, Donatella Firmani, Paolo Merialdo, Tommaso TeofiliSIGMOD 2022 · 47 citations
- Overlooked Factors in Concept-Based Explanations: Dataset Choice, Concept Learnability, and Human CapabilityVikram V. Ramaswamy, Sunnie S. Y. Kim, Ruth Fong, Olga RussakovskyCVPR 2023
