Manifold Integrated Gradients: Riemannian Geometry for Feature Attribution
Eslam Zaher, Maciej Trzaskowski, Quan Nguyen, Fred Roosta
Abstract
In this paper, we dive into the reliability concerns of Integrated Gradients (IG), a prevalent feature attribution method for black-box deep learning models. We particularly address two predominant challenges associated with IG: the generation of noisy feature visualizations for vision models and the vulnerability to adversarial attributional attacks. Our approach involves an adaptation of path-based feature attribution, aligning the path of attribution more closely to the intrinsic geometry of the data manifold. Our experiments utilise deep generative models applied to several real-world image datasets. They demonstrate that IG along the geodesics conforms to the curved geometry of the Riemannian data manifold, generating more perceptually intuitive explanations and, subsequently, substantially increasing robustness to targeted attributional attacks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5bdd6f53-8a9f-452d-b06f-4cfae6ceafe1Cited by top-tier papers6
- AdaptGrad: Adaptive Sampling to Reduce NoiseLinjiang Zhou, Chao Ma, Zepeng Wang, Libing Wu et al.NeurIPS 2025 · 3 citations
- Auditing Sybil: Explaining Deep Lung Cancer Risk Prediction Through Generative Interventional AttributionsBartlomiej Sobieski, Jakub Grzywaczewski, Karol Dobiczek, Mateusz Wójcik et al.ICML 2026 · 2 citations
- Spectral Integrated Gradients for Coarse-to-Fine Feature AttributionSoyeon Kim, Seongwoo Lim, Kyowoon Lee, Jaesik ChoiKDD 2026 · 2 citations
- Manifold-Aligned Guided Integrated Gradients for Reliable Feature AttributionSoyeon Kim, Seongwoo Lim, Kyowoon Lee, Jaesik ChoiICML 2026 · 2 citations
- Counterfactual Explanations on Robust Perceptual GeodesicsEslam Zaher, Dr Maciej Trzaskowski, Quan Nguyen, Fred RoostaICLR 2026 · 2 citations
Builds on12
- Concise Explanations of Neural Networks using Adversarial TrainingPrasad Chalasani, Jiefeng Chen, Amrita Roy Chowdhury, Xi Wu et al.ICML 2020 · 148 citations
- Mixed-curvature Variational AutoencodersOndrej Skopek, Octavian-Eugen Ganea, Gary BécigneulICLR 2020 · 122 citations
- A Rigorous Study of Integrated Gradients Method and Extensions to Internal Neuron AttributionsDaniel Lundström, Tianjian Huang, Meisam RazaviyaynICML 2022 · 85 citations
- Do Input Gradients Highlight Discriminative Features?Harshay Shah, Prateek Jain, Praneeth NetrapalliNeurIPS 2021 · 74 citations
- Which Models have Perceptually-Aligned Gradients? An Explanation via Off-Manifold RobustnessSuraj Srinivas, Sebastian Bordt, Himabindu LakkarajuNeurIPS 2023 · 24 citations
Related papers
- Guided Integrated Gradients: An Adaptive Path Method for Removing NoiseAndrei Kapishnikov, Subhashini Venugopalan, Besim Avci, Ben Wedin et al.CVPR 2021
- Smoothed Geometry for Robust AttributionZifan Wang, Haofan Wang, Shakul Ramkumar, Piotr Mardziel et al.NeurIPS 2020 · 67 citations
- Denoising Diffusion Path: Attribution Noise Reduction with An Auxiliary Diffusion ModelYiming Lei, Zilong Li, Junping Zhang, Hongming ShanNeurIPS 2024 · 9 citations
- Beyond Single Path Integrated Gradients for Reliable Input Attribution via Randomized Path SamplingGiyoung Jeon, Haedong Jeong, Jaesik ChoiICCV 2023 · 3 citations
- Local Path Integration for AttributionPeiyu Yang, Naveed Akhtar, Zeyi Wen, Ajmal MianAAAI 2023 · 16 citations
