Variable Importance in High-Dimensional Settings Requires Grouping
Ahmad Chamma, Bertrand Thirion, Denis A. Engemann
Abstract
Explaining the decision process of machine learning algorithms is nowadays crucial for both a model's performance enhancement and human comprehension. This can be achieved by assessing the variable importance of single variables, even for high-capacity non-linear methods, e.g. Deep Neural Networks (DNNs). While only removalbased approaches, such as Permutation Importance (PI), can bring statistical validity, they return misleading results when variables are correlated. Conditional Permutation Importance (CPI) bypasses PI's limitations in such cases. However, in high-dimensional settings, where high correlations between the variables cancel their conditional importance, the use of CPI as well as other methods leads to unreliable results, besides prohibitive computation costs. Grouping variables statistically via clustering or some prior knowledge gains some power back and leads to better interpretations. In this work, we introduce BCPI (Block-Based Conditional Permutation Importance), a new generic framework for variable importance computation with statistical guarantees handling both single and group cases. Furthermore, as handling groups with high cardinality (such as a set of observations of a given modality) are both time-consuming and resource-intensive, we also introduce a new stacking approach extending the DNN architecture with sub-linear layers adapted to the group structure. We show that the ensuing approach extended with stacking controls the type-I error even with highly-correlated groups and shows top accuracy across benchmarks. Furthermore, we perform a real-world data analysis in a large-scale medical dataset where we aim to show the consistency between our results and the literature for a biomarker prediction.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Flow-Disentangled Feature ImportanceXingshu Chen, Yifeng Guo, Jin-Hong DuICLR 2026
- Measuring Variable Importance in Heterogeneous Treatment Effects with ConfidenceJoseph Paillard, Angel David Reyero Lobo, Vitaliy Kolodyazhniy, Bertrand Thirion et al.ICML 2025
Builds on3
- Understanding Global Feature Contributions With Additive Importance MeasuresIan Covert, Scott M. Lundberg, Su-In LeeNeurIPS 2020 · 476 citations
- Statistically Valid Variable Importance Assessment through Conditional PermutationsAhmad Chamma, Denis A. Engemann, Bertrand ThirionNeurIPS 2023 · 23 citations
- Lazy Estimation of Variable Importance for Large Neural NetworksYue Gao, Abby Stevens, Garvesh Raskutti, Rebecca WillettICML 2022 · 7 citations
Related papers
- Testing Conditional Mean Independence Using Generative Neural NetworksYi Zhang, Linjun Huang, Yun Yang, Xiaofeng ShaoICML 2025
- Covered Information Disentanglement: Model Transparency via Unbiased Permutation ImportanceJoão P. B. Pereira, Erik S. G. Stroes, Aeilko H. Zwinderman, Evgeni LevinAAAI 2022 · 17 citations
- Aggregate Models, Not Explanations: Improving Feature Importance EstimationJoseph Paillard, Angel REYERO LOBO, Denis-Alexander Engemann, Thirion BertrandICML 2026 · 1 citation
- Permutation-Based Hypothesis Testing for Neural NetworksFrancesca Mandel, Ian BarnettAAAI 2024 · 6 citations
- Explainability as statistical inferenceHugo Henri Joseph Senetaire, Damien Garreau, Jes Frellsen, Pierre-Alexandre MatteiICML 2023 · 4 citations
