SOInter: A Novel Deep Energy-Based Interpretation Method for Explaining Structured Output Models
Seyyede Fatemeh Seyyedsalehi, Mahdieh Soleymani Baghshah, Hamid R. Rabiee
Abstract
We propose a novel interpretation technique to explain the behavior of structured output models, which learn mappings between an input vector to a set of output variables simultaneously. Because of the complex relationship between the computational path of output variables in structured models, a feature can affect the value of output through other ones. We focus on one of the outputs as the target and try to find the most important features utilized by the structured model to decide on the target in each locality of the input space. In this paper, we assume an arbitrary structured output model is available as a black box and argue how considering the correlations between output variables can improve the explanation performance. The goal is to train a function as an interpreter for the target output variable over the input space. We introduce an energy-based training process for the interpreter function, which effectively considers the structural information incorporated into the model to be explained. The effectiveness of the proposed method is confirmed using a variety of simulated and real data sets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Related papers
- Visual Neural Decomposition to Explain Multivariate Data SetsJohannes Knittel, Andrés Lalama, Steffen Koch, Thomas ErtlIEEE VIS 2020 · 14 citations
- Unsupervised Causal Binary Concepts Discovery with VAE for Black-Box Model ExplanationThien Q. Tran, Kazuto Fukuchi, Youhei Akimoto, Jun SakumaAAAI 2022 · 11 citations
- An Additive Instance-Wise Approach to Multi-class Model InterpretationVy Vo, Van Nguyen, Trung Le, Quan Hung Tran et al.ICLR 2023
- Towards Interpretation of Pairwise LearningMengdi Huai, Di Wang, Chenglin Miao, Aidong ZhangAAAI 2020 · 8 citations
- A Framework to Learn with InterpretationJayneel Parekh, Pavlo Mozharovskyi, Florence d'Alché-BucNeurIPS 2021 · 35 citations
