SAND: One-Shot Feature Selection with Additive Noise Distortion
Pedram Pad, Hadi Hammoud, Mohamad Dia, Nadim Maamari, Liza Andrea Dunbar
摘要
Feature selection is a critical step in data-driven applications, reducing input dimensionality to enhance learning accuracy, computational efficiency, and interpretability. Existing state-of-theart methods often require post-selection retraining and extensive hyperparameter tuning, complicating their adoption. We introduce a novel, non-intrusive feature selection layer that, given a target feature count k, automatically identifies and selects the k most informative features during neural network training. Our method is uniquely simple, requiring no alterations to the loss function, network architecture, or post-selection retraining. The layer is mathematically elegant and can be fully described by: xi = a i x i + (1 -a i )z i where x i is the input feature, xi the output, z i a Gaussian noise, and a i trainable gain such that i a 2 i = k. This formulation induces an automatic clustering effect, driving k of the a i gains to 1 (selecting informative features) and the rest to 0 (discarding redundant ones) via weighted noise distortion and gain normalization. Despite its extreme simplicity, our method achieves competitive performance on standard benchmark datasets and a novel real-world dataset, often matching or exceeding existing approaches without requiring hyperparameter search for k or retraining. Theoretical analysis in the context of linear regression further validates its efficacy. Our work demonstrates that simplicity and performance are not mutually exclusive, offering a powerful yet straightforward tool for feature selection in machine learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Feature Selection using Stochastic GatesYutaro Yamada, Ofir Lindenbaum, Sahand Negahban, Yuval KlugerICML 2020 · 被引用 39 次
- Where to Pay Attention in Sparse Training for Feature Selection?Ghada Sokar, Zahra Atashgahi, Mykola Pechenizkiy, Decebal Constantin MocanuNeurIPS 2022 · 被引用 25 次
- Sequential Attention for Feature SelectionTaisuke Yasuda, Mohammad Hossein Bateni, Lin Chen, Matthew Fahrbach 等ICLR 2023 · 被引用 2 次
相关 Paper
- Overcoming Simplicity Bias in Deep Networks using a Feature SieveRishabh Tiwari, Pradeep ShenoyICML 2023 · 被引用 32 次
- Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature EfficiencyNaoki Nishikawa, Rei Higuchi, Taiji SuzukiNeurIPS 2025 · 被引用 2 次
- Fractal Autoencoders for Feature SelectionXinxing Wu, Qiang ChengAAAI 2021 · 被引用 33 次
- Pay Attention to Features, Transfer Learn Faster CNNsKafeng Wang, Xitong Gao, Yiren Zhao, Xingjian Li 等ICLR 2020 · 被引用 83 次
- NEAR: A Training-Free Pre-Estimator of Machine Learning Model PerformanceRaphael T. Husistein, Markus Reiher, Marco EckhoffICLR 2025
