Beyond Trivial Counterfactual Explanations with Diverse Valuable Explanations
Pau Rodríguez, Massimo Caccia, Alexandre Lacoste, Lee Zamparo, Issam H. Laradji, Laurent Charlin, David Vázquez
Abstract
Explainability for machine learning models has gained considerable attention within the research community given the importance of deploying more reliable machinelearning systems. In computer vision applications, generative counterfactual methods indicate how to perturb a model's input to change its prediction, providing details about the model's decision-making. Current methods tend to generate trivial counterfactuals about a model's decisions, as they often suggest to exaggerate or remove the presence of the attribute being classified. For the machine learning practitioner, these types of counterfactuals offer little value, since they provide no new information about undesired model or data biases. In this work, we identify the problem of trivial counterfactual generation and we propose DiVE to alleviate it. DiVE learns a perturbation in a disentangled latent space that is constrained using a diversity-enforcing loss to uncover multiple valuable explanations about the model's prediction. Further, we introduce a mechanism to prevent the model from producing trivial explanations. Experiments on CelebA and Synbols demonstrate that our model improves the success rate of producing high-quality valuable explanations when compared to previous state-of-the-art methods. Code is available at https://github.com/ElementAI/ beyond-trivial-explanations .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 425c5647-7579-4074-99b1-8cb2e91c10a4Cited by top-tier papers13
- Cycle-Consistent Counterfactuals by Latent TransformationsSaeed Khorram, Fuxin LiCVPR 2022 · 27 citations
- Counterfactual Image EditingYushu Pan, Elias BareinboimICML 2024 · 19 citations
- CounterNet: End-to-End Training of Prediction Aware Counterfactual ExplanationsHangzhi Guo, Thanh Hong Nguyen, Amulya YadavKDD 2023 · 12 citations
- On the explainable properties of 1-Lipschitz Neural Networks: An Optimal Transport PerspectiveMathieu Serrurier, Franck Mamalet, Thomas Fel, Louis Béthune et al.NeurIPS 2023 · 11 citations
- Interpretable Knowledge Tracing via Response Influence-based Counterfactual ReasoningJiajun Cui, Minghe Yu, Bo Jiang, Aimin Zhou et al.ICDE 2024 · 11 citations
Builds on5
- Weakly-Supervised Disentanglement Without CompromisesFrancesco Locatello, Ben Poole, Gunnar Rätsch, Bernhard Schölkopf et al.ICML 2020 · 361 citations
- Deep Structural Causal Models for Tractable Counterfactual InferenceNick Pawlowski, Daniel Coelho de Castro, Ben GlockerNeurIPS 2020 · 353 citations
- Counterfactual Generative NetworksAxel Sauer, Andreas GeigerICLR 2021 · 145 citations
- Explanation by Progressive ExaggerationSumedha Singla, Brian Pollack, Junxiang Chen, Kayhan BatmanghelichICLR 2020 · 116 citations
- Synbols: Probing Learning Algorithms with Synthetic DatasetsAlexandre Lacoste, Pau Rodríguez López, Frederic Branchaud-Charron, Parmida Atighehchian et al.NeurIPS 2020 · 14 citations
Related papers
- Constructing Fair Latent Space for Intersection of Fairness and ExplainabilityHyungjun Joo, Hyeonggeun Han, Sehwan Kim, Sangwoo Hong et al.AAAI 2025 · 2 citations
- DECE: Decision Explorer with Counterfactual Explanations for Machine Learning ModelsFurui Cheng, Yao Ming, Huamin QuIEEE VIS 2020 · 118 citations
- LeapFactual: Reliable Visual Counterfactual Explanation Using Conditional Flow MatchingZhuo Cao, Xuan Zhao, Lena Krieger, Hanno Scharr et al.NeurIPS 2025 · 5 citations
- DISSECT: Disentangled Simultaneous Explanations via Concept TraversalsAsma Ghandeharioun, Been Kim, Chun-Liang Li, Brendan Jou et al.ICLR 2022 · 58 citations
- OCTET: Object-aware Counterfactual ExplanationsMehdi Zemni, Mickaël Chen, Éloi Zablocki, Hédi Ben-Younes et al.CVPR 2023
