Feature Importance Ranking for Deep Learning
Maksymilian Wojtas, Ke Chen
Abstract
Feature importance ranking has become a powerful tool for explainable AI. However, its nature of combinatorial optimization poses a great challenge for deep learning. In this paper, we propose a novel dual-net architecture consisting of operator and selector for discovery of an optimal feature subset of a fixed size and ranking the importance of those features in the optimal subset simultaneously. During learning, the operator is trained for a supervised learning task via optimal feature subset candidates generated by the selector that learns predicting the learning performance of the operator working on different optimal subset candidates. We develop an alternate learning algorithm that trains two nets jointly and incorporates a stochastic local search procedure into learning to address the combinatorial optimization challenge. In deployment, the selector generates an optimal feature subset and ranks feature importance, while the operator makes predictions based on the optimal subset for test data. A thorough evaluation on synthetic, benchmark and real data sets suggests that our approach outperforms several state-of-the-art feature importance ranking and supervised feature selection methods. (Our source code is available: https://github.com/maksym33/FeatureImportanceDL ) 34th Conference on Neural Information Processing Systems (NeurIPS 2020),
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 73e71240-a079-4729-8fca-5b6b05ab1919Cited by top-tier papers10
- The Out-of-Distribution Problem in Explainability and Search Methods for Feature Importance ExplanationsPeter Hase, Harry Xie, Mohit BansalNeurIPS 2021 · 121 citations
- Accelerating Shapley Explanation via Contributive Cooperator SelectionGuanchu Wang, Yu-Neng Chuang, Mengnan Du, Fan Yang et al.ICML 2022 · 25 citations
- Efficient Top-K Feature Selection Using Coordinate Descent MethodLei Xu, Rong Wang, Feiping Nie, Xuelong LiAAAI 2023 · 20 citations
- The Modality Focusing Hypothesis: Towards Understanding Crossmodal Knowledge DistillationZihui Xue, Zhengqi Gao, Sucheng Ren, Hang ZhaoICLR 2023 · 12 citations
- Permutation-Based Hypothesis Testing for Neural NetworksFrancesca Mandel, Ian BarnettAAAI 2024 · 6 citations
Related papers
- Active feature acquisition via explainability-driven rankingOsman Berke Güney, Ketan Suhaas Saichandran, Karim Elzokm, Ziming Zhang et al.ICML 2025
- Automatic Feature Selection By One-Shot Neural Architecture Search In Recommendation SystemsHe Wei, Yuekui Yang, Haiyang Wu, Yangyang Tang et al.WWW 2023 · 5 citations
- AutoAL: Automated Active Learning with Differentiable Query Strategy SearchYifeng Wang, Xueying Zhan, Siyu HuangICML 2025
- Explainable Neural Networks with Guarantee: A Sparse Estimation ApproachAntoine Ledent, Peng LiuAAAI 2025 · 1 citation
- Learning Deep Attribution Priors Based On Prior KnowledgeEthan Weinberger, Joseph D. Janizek, Su-In LeeNeurIPS 2020 · 27 citations
