Feature Selection using Stochastic Gates
Yutaro Yamada, Ofir Lindenbaum, Sahand Negahban, Yuval Kluger
摘要
Feature selection problems have been extensively studied in the setting of linear estimation (e.g. LASSO), but less emphasis has been placed on feature selection for non-linear functions. In this study, we propose a method for feature selection in neural network estimation problems. The new procedure is based on probabilistic relaxation of the 0 norm of features, or the count of the number of selected features. Our 0 -based regularization relies on a continuous relaxation of the Bernoulli distribution; such relaxation allows our model to learn the parameters of the approximate Bernoulli distributions via gradient descent. The proposed framework simultaneously learns either a nonlinear regression or classification function while selecting a small subset of features. We provide an information-theoretic justification for incorporating Bernoulli distribution into feature selection. Furthermore, we evaluate our method using synthetic and real-life data to demonstrate that our approach outperforms other commonly used methods in both predictive performance and feature selection.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper50
- Learning to Maximize Mutual Information for Dynamic Feature SelectionIan Connick Covert, Wei Qiu, Mingyu Lu, Nayoon Kim 等ICML 2023 · 被引用 67 次
- TuneTables: Context Optimization for Scalable Prior-Data Fitted NetworksBenjamin Feuer, Robin Schirrmeister, Valeriia Cherepanova, Chinmay Hegde 等NeurIPS 2024 · 被引用 57 次
- FedSDG-FS: Efficient and Secure Feature Selection for Vertical Federated LearningAnran Li, Hongyi Peng, Lan Zhang, Jiahui Huang 等INFOCOM 2023 · 被引用 50 次
- Locally Sparse Neural Networks for Tabular Biomedical DataJunchen Yang, Ofir Lindenbaum, Yuval KlugerICML 2022 · 被引用 45 次
- Differentiable Unsupervised Feature Selection based on a Gated LaplacianOfir Lindenbaum, Uri Shaham, Erez Peterfreund, Jonathan Svirsky 等NeurIPS 2021 · 被引用 38 次
相关 Paper
- Discovering Features with Synergistic Interactions in Multiple ViewsChohee Kim, Mihaela van der Schaar, Changhee LeeICML 2024 · 被引用 4 次
- Few-shot Learning for Feature Selection with Hilbert-Schmidt Independence CriterionAtsutoshi Kumagai, Tomoharu Iwata, Yasutoshi Ida, Yasuhiro FujiwaraNeurIPS 2022 · 被引用 13 次
- Training Binary Neural Networks using the Bayesian Learning RuleXiangming Meng, Roman Bachmann, Mohammad Emtiyaz KhanICML 2020 · 被引用 47 次
- Safe screening rules for L0-regression from Perspective RelaxationsAlper Atamtürk, Andrés GómezICML 2020 · 被引用 12 次
- Deep Learning meets Nonparametric Regression: Are Weight-Decayed DNNs Locally Adaptive?Kaiqi Zhang, Yu-Xiang WangICLR 2023 · 被引用 3 次
