Explainable Models with Consistent Interpretations
Vipin Pillai, Hamed Pirsiavash
摘要
Given the widespread deployment of black box deep neural networks in computer vision applications, the interpretability aspect of these black box systems has recently gained traction. Various methods have been proposed to explain the results of such deep neural networks. However, some recent works have shown that such explanation methods are biased and do not produce consistent interpretations. Hence, rather than introducing a novel explanation method, we learn models that are encouraged to be interpretable given an explanation method. We use Grad-CAM as the explanation algorithm and encourage the network to learn consistent interpretations along with maximizing the log-likelihood of the correct class. We show that our method outperforms the baseline on the pointing game evaluation on ImageNet and MS-COCO datasets respectively. We also introduce new evaluation metrics that penalize the saliency map if it lies outside the ground truth bounding box or segmentation mask, and show that our method outperforms the baseline on these metrics as well. Moreover, our model trained with interpretation consistency generalizes to other explanation algorithms on all the evaluation metrics. The code and models are publicly available.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Learning Support and Trivial Prototypes for Interpretable Image ClassificationChong Wang, Yuyuan Liu, Yuanhong Chen, Fengbei Liu 等ICCV 2023 · 被引用 50 次
- Studying How to Efficiently and Effectively Guide Models with ExplanationsSukrut Rao, Moritz Böhle, Amin Parchami-Araghi, Bernt SchieleICCV 2023 · 被引用 22 次
- Saliency-Aware Neural Architecture SearchRamtin Hosseini, Pengtao XieNeurIPS 2022 · 被引用 16 次
- Consistent Explanations by Contrastive LearningVipin Pillai, Soroush Abbasi Koohpayegani, Ashley Ouligian, Dennis Fong 等CVPR 2022 · 被引用 15 次
- B-cosification: Transforming Deep Neural Networks to be Inherently InterpretableShreyash Arya, Sukrut Rao, Moritz Böhle, Bernt SchieleNeurIPS 2024 · 被引用 14 次
它引用的顶会 Paper4
- Fooling Network Interpretation in Image ClassificationAkshayvarun Subramanya, Vipin Pillai, Hamed PirsiavashICCV 2019 · 被引用 68 次
- Self-Supervised Equivariant Attention Mechanism for Weakly Supervised Semantic SegmentationYude Wang, Jie Zhang, Meina Kan, Shiguang Shan 等CVPR 2020
- Don't Judge an Object by Its Context: Learning to Overcome Contextual BiasKrishna Kumar Singh, Dhruv Mahajan, Kristen Grauman, Yong Jae Lee 等CVPR 2020
- Momentum Contrast for Unsupervised Visual Representation LearningKaiming He, Haoqi Fan, Yuxin Wu, Saining Xie 等CVPR 2020
相关 Paper
- What You See is What You Classify: Black Box AttributionsSteven Stalder, Nathanaël Perraudin, Radhakrishna Achanta, Fernando Pérez-Cruz 等NeurIPS 2022 · 被引用 15 次
- Learning Global Transparent Models consistent with Local Contrastive ExplanationsTejaswini Pedapati, Avinash Balakrishnan, Karthikeyan Shanmugam, Amit DhurandharNeurIPS 2020 · 被引用 35 次
- LICO: Explainable Models with Language-Image COnsistencyYiming Lei, Zilong Li, Yangyang Li, Junping Zhang 等NeurIPS 2023 · 被引用 12 次
- A Novel Visual Interpretability for Deep Neural Networks by Optimizing Activation Maps with PerturbationQing-Long Zhang, Lu Rao, Yubin YangAAAI 2021 · 被引用 26 次
- Eye into AI: Evaluating the Interpretability of Explainable AI Techniques through a Game with a PurposeKatelyn Morrison, Mayank Jain, Jessica Hammer, Adam PererCSCW 2023 · 被引用 11 次
