Tackling the XAI Disagreement Problem with Adaptive Feature Grouping
Gabriel Laberge, Ola Ahmad
Abstract
Post-hoc explanations aim at understanding which input features (or groups thereof) are the most impactful toward certain model decisions. Many such methods have been proposed (ArchAttribute, Occlusion, SHAP, RISE, LIME, Integrated Gradient) and it is hard for practitioners to understand the differences between them. Even worse, faithfulness metrics, often used to quantitatively compare explanation methods, also exhibit inconsistencies. To address these issues, recent work has unified explanation methods through the lens of Functional Decomposition. We extend such work to scenarios where input features are partitioned into groups (e.g. pixel patches) and prove that disagreements between explanation methods and faithfulness metrics are caused by between-group interactions. Crucially, getting rid of between-group interactions leads to a single explanation that is optimal according to all faithfulness metrics. We finally show how to reduce the disagreements by grouping features on tabular/image data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2e35c313-75a3-4f02-99f3-ecb569c322dcBuilds on7
- The Many Shapley Values for Model ExplanationMukund Sundararajan, Amir NajmiICML 2020 · 799 citations
- Sanity Checks for Saliency MetricsRichard Tomsett, Dan Harborne, Supriyo Chakraborty, Prudhvi Gurram et al.AAAI 2020 · 204 citations
- How does This Interaction Affect Me? Interpretable Attribution for Feature InteractionsMichael Tsang, Sirisha Rambhatla, Yan LiuNeurIPS 2020 · 109 citations
- On Gradient-like Explanation under a Black-box Setting: When Black-box Explanations Become as Good as White-boxYi Cai, Gerhard WunderICML 2024 · 4 citations
- CRAFT: Concept Recursive Activation FacTorization for ExplainabilityThomas Fel, Agustin Martin Picard, Louis Béthune, Thibaut Boissin et al.CVPR 2023
Related papers
- Which Explanation Should I Choose? A Function Approximation Perspective to Characterizing Post Hoc ExplanationsTessa Han, Suraj Srinivas, Himabindu LakkarajuNeurIPS 2022 · 126 citations
- Regional Explanations: Bridging Local and Global Variable ImportanceSalim I. Amoukou, Nicolas J.-B. BrunelNeurIPS 2025
- Provably Better Explanations with Optimized Aggregation of Feature AttributionsThomas Decker, Ananta R. Bhattarai, Jindong Gu, Volker Tresp et al.ICML 2024 · 7 citations
- Towards Attributions of Input Variables in a CoalitionXinhao Zheng, Huiqi Deng, Quanshi ZhangICML 2025
- Succinct Interaction-Aware ExplanationsSascha Xu, Joscha Cüppers, Jilles VreekenKDD 2025
