Explaining Local, Global, And Higher-Order Interactions In Deep Learning
Samuel Lerman, Charles Venuto, Henry A. Kautz, Chenliang Xu
摘要
We present a simple yet highly generalizable method for explaining interacting parts within a neural network’s reasoning process. First, we design an algorithm based on cross derivatives for computing statistical interaction effects between individual features, which is generalized to both 2-way and higher-order (3-way or more) interactions. We present results side by side with a weight-based attribution technique, corroborating that cross derivatives are a superior metric for both 2-way and higher-order interaction detection. Moreover, we extend the use of cross derivatives as an explanatory device in neural networks to the computer vision setting by expanding Grad-CAM, a popular gradient-based explanatory tool for CNNs, to the higher order. While Grad-CAM can only explain the importance of individual objects in images, our method, which we call Taylor-CAM, can explain a neural network’s relational reasoning across multiple objects. We show the success of our explanations both qualitatively and quantitatively, including with a user study. We will release all code as a tool package to facilitate explainable deep learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- HorNet: Efficient High-Order Spatial Interactions with Recursive Gated ConvolutionsYongming Rao, Wenliang Zhao, Yansong Tang, Jie Zhou 等NeurIPS 2022 · 被引用 422 次
- SmoothHess: ReLU Network Feature Interactions via Stein's LemmaMax Torop, Aria Masoomi, Davin Hill, Kivanç Köse 等NeurIPS 2023 · 被引用 9 次
- Synthesizing Precise Static Analyzers for Automatic DifferentiationJacob Laurel, Siyuan Brant Qian, Gagandeep Singh, Sasa MisailovicOOPSLA 2023 · 被引用 8 次
它引用的顶会 Paper3
- How does This Interaction Affect Me? Interpretable Attribution for Feature InteractionsMichael Tsang, Sirisha Rambhatla, Yan LiuNeurIPS 2020 · 被引用 109 次
- Feature Interaction Interpretability: A Case for Explaining Ad-Recommendation Systems via Neural Interaction DetectionMichael Tsang, Dehua Cheng, Hanpeng Liu, Xue Feng 等ICLR 2020 · 被引用 71 次
- Don't Judge an Object by Its Context: Learning to Overcome Contextual BiasKrishna Kumar Singh, Dhruv Mahajan, Kristen Grauman, Yong Jae Lee 等CVPR 2020
相关 Paper
- Fooling Network Interpretation in Image ClassificationAkshayvarun Subramanya, Vipin Pillai, Hamed PirsiavashICCV 2019 · 被引用 68 次
- Empowering CAM-Based Methods with Capability to Generate Fine-Grained and High-Faithfulness ExplanationsChangqing Qiu, Fusheng Jin, Yining ZhangAAAI 2024 · 被引用 11 次
- Identifying Important Group of Pixels using InteractionsKosuke Sumiyasu, Kazuhiko Kawamoto, Hiroshi KeraCVPR 2024
- Transformer Interpretability Beyond Attention VisualizationHila Chefer, Shir Gur, Lior WolfCVPR 2021
- Relevance-CAM: Your Model Already Knows Where To LookJeong Ryong Lee, Sewon Kim, Inyong Park, Taejoon Eo 等CVPR 2021
