Consistent feature selection for analytic deep neural networks
Vu C. Dinh, Lam Si Tung Ho
摘要
One of the most important steps toward interpretability and explainability of neural network models is feature selection, which aims to identify the subset of relevant features. Theoretical results in the field have mostly focused on the prediction aspect of the problem with virtually no work on feature selection consistency for deep neural networks due to the model's severe nonlinearity and unidentifiability. This lack of theoretical foundation casts doubt on the applicability of deep learning to contexts where correct interpretations of the features play a central role. In this work, we investigate the problem of feature selection for analytic deep networks. We prove that for a wide class of networks, including deep feed-forward neural networks, convolutional neural networks, and a major sub-class of residual neural networks, the Adaptive Group Lasso selection procedure with Group Lasso as the base estimator is selection-consistent. The work provides further evidence that Group Lasso might be inefficient for feature selection with neural networks and advocates the use of Adaptive Group Lasso over the popular Group Lasso.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Multimodal Dynamics: Dynamical Fusion for Trustworthy Multimodal ClassificationZongbo Han, Fan Yang, Junzhou Huang, Changqing Zhang 等CVPR 2022 · 被引用 149 次
- Neural graphical modelling in continuous-time: consistency guarantees and algorithmsAlexis Bellot, Kim Branson, Mihaela van der SchaarICLR 2022 · 被引用 57 次
- LESS-VFL: Communication-Efficient Feature Selection for Vertical Federated LearningTimothy Castiglia, Yi Zhou, Shiqiang Wang, Swanand Kadhe 等ICML 2023 · 被引用 33 次
- Why Lottery Ticket Wins? A Theoretical Perspective of Sample Complexity on Sparse Neural NetworksShuai Zhang, Meng Wang, Sijia Liu, Pin-Yu Chen 等NeurIPS 2021 · 被引用 28 次
- Bilevel Network Learning via Hierarchically Structured SparsityJiayi Fan, Jingyuan Yang, Shuangge Ma, Mengyun WuNeurIPS 2025 · 被引用 1 次
相关 Paper
- Global Optimality Beyond Two Layers: Training Deep ReLU Networks via Convex ProgramsTolga Ergen, Mert PilanciICML 2021 · 被引用 35 次
- The Contextual Lasso: Sparse Linear Models via Deep Neural NetworksRyan Thompson, Amir Dezfouli, Robert KohnNeurIPS 2023 · 被引用 8 次
- Algorithmic stability and generalization of an unsupervised feature selection algorithmXinxing Wu, Qiang ChengNeurIPS 2021 · 被引用 13 次
- Theoretical Characterisation of the Gauss Newton Conditioning in Neural NetworksJim Zhao, Sidak Pal Singh, Aurélien LucchiNeurIPS 2024 · 被引用 7 次
- Architectural Adversarial Robustness: The Case for Deep PursuitGeorge Cazenavette, Calvin Murdock, Simon LuceyCVPR 2021
