Towards Fine-Grained Interpretability: Counterfactual Explanations for Misclassification with Saliency Partition
Lintong Zhang, Kang Yin, Seong-Whan Lee
Abstract
Attribution-based explanation techniques capture key patterns to enhance visual interpretability; however, these patterns often lack the granularity needed for insight in fine-grained tasks, particularly in cases of model misclassification, where explanations may be insufficiently detailed. To address this limitation, we propose a fine-grained counterfactual explanation framework that generates both object-level and part-level interpretability, addressing two fundamental questions: (1) which fine-grained features contribute to model misclassification, and (2) where dominant local features influence counterfactual adjustments. Our approach yields explainable counterfactuals in a non-generative manner by quantifying similarity and weighting component contributions within regions of interest between correctly classified and misclassified samples. Furthermore, we introduce a saliency partition module grounded in Shapley value contributions, isolating features with region-specific relevance. Extensive experiments demonstrate the superiority of our approach in capturing more granular, intuitively meaningful regions, surpassing fine-grained methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 84a1967a-165f-4a7f-8f74-64d116d1d090Cited by top-tier papers1
Ask how each one uses itBuilds on11
- Understanding Deep Networks via Extremal Perturbations and Smooth MasksRuth Fong, Mandela Patrick, Andrea VedaldiICCV 2019 · 480 citations
- Relative Attributing Propagation: Interpreting the Comparative Contributions of Individual Units in Deep Neural NetworksWoo-Jeoung Nam, Shir Gur, Jaesik Choi, Lior Wolf et al.AAAI 2020 · 109 citations
- CoCoX: Generating Conceptual and Counterfactual Explanations via Fault-LinesArjun R. Akula, Shuai Wang, Song-Chun ZhuAAAI 2020 · 102 citations
- Stochastic Partial Swap: Enhanced Model Generalization and Interpretability for Fine-grained RecognitionShaoli Huang, Xinchao Wang, Dacheng TaoICCV 2021 · 46 citations
- Keep CALM and Improve Visual Feature AttributionJae-Myung Kim, Junsuk Choe, Zeynep Akata, Seong Joon OhICCV 2021 · 22 citations
Related papers
- Explaining Object Detectors via Collective Contribution of PixelsToshinori Yamauchi, Hiroshi Kera, Kazuhiko KawamotoCVPR 2026
- Interpretable and Accurate Fine-grained Recognition via Region GroupingZixuan Huang, Yin LiCVPR 2020
- Towards Attributions of Input Variables in a CoalitionXinhao Zheng, Huiqi Deng, Quanshi ZhangICML 2025
- The Many Shapley Values for Model ExplanationMukund Sundararajan, Amir NajmiICML 2020 · 799 citations
- ShapeX: Shapelet-Driven Post Hoc Explanations for Time Series Classification ModelsBosong Huang, Ming Jin, Yuxuan Liang, Johan Barthelemy et al.NeurIPS 2025 · 8 citations
