Locally Sparse Neural Networks for Tabular Biomedical Data
Junchen Yang, Ofir Lindenbaum, Yuval Kluger
Abstract
Tabular datasets with low-sample-size or many variables are prevalent in biomedicine. Practitioners in this domain prefer linear or tree-based models over neural networks since the latter are harder to interpret and tend to overfit when applied to tabular datasets. To address these neural networks' shortcomings, we propose an intrinsically interpretable network for heterogeneous biomedical data. We design a locally sparse neural network where the local sparsity is learned to identify the subset of most relevant features for each sample. This sample-specific sparsity is predicted via a gating network, which is trained in tandem with the prediction network. By forcing the model to select a subset of the most informative features for each sample, we reduce model overfitting in low-sample-size data and obtain an interpretable model. We demonstrate that our method outperforms state-of-the-art models when applied to synthetic or real-world biomedical datasets using extensive experiments. Furthermore, the proposed framework dramatically outperforms existing schemes when evaluating its interpretability capabilities. Finally, we demonstrate the applicability of our model to two important biomedical tasks: survival analysis and marker gene identification.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bba5ac11-d599-4a73-822e-091acbfa24cfCited by top-tier papers16
- Explaining Time Series via Contrastive and Locally Sparse PerturbationsZichuan Liu, Yingying Zhang, Tianchun Wang, Zefan Wang et al.ICLR 2024 · 26 citations
- ProtoGate: Prototype-based Neural Networks with Global-to-local Feature Selection for Tabular Biomedical DataXiangjian Jiang, Andrei Margeloiu, Nikola Simidjievski, Mateja JamnikICML 2024 · 23 citations
- TabSTAR: A Tabular Foundation Model for Tabular Data with Text FieldsAlan Arazi, Eilam Shapira, Roi ReichartNeurIPS 2025 · 20 citations
- Interpretable Deep Clustering for Tabular DataJonathan Svirsky, Ofir LindenbaumICML 2024 · 19 citations
- Transfer Learning with Deep Tabular ModelsRoman Levin, Valeriia Cherepanova, Avi Schwarzschild, Arpit Bansal et al.ICLR 2023 · 18 citations
Builds on4
- TabNet: Attentive Interpretable Tabular LearningSercan Ö. Arik, Tomas PfisterAAAI 2021 · 2,148 citations
- Scaling Vision with Sparse Mixture of ExpertsCarlos Riquelme, Joan Puigcerver, Basil Mustafa, Maxim Neumann et al.NeurIPS 2021 · 1,213 citations
- Feature Selection using Stochastic GatesYutaro Yamada, Ofir Lindenbaum, Sahand Negahban, Yuval KlugerICML 2020 · 39 citations
- Differentiable Unsupervised Feature Selection based on a Gated LaplacianOfir Lindenbaum, Uri Shaham, Erez Peterfreund, Jonathan Svirsky et al.NeurIPS 2021 · 38 citations
Related papers
- Weight Predictor Network with Feature Selection for Small Sample Tabular Biomedical DataAndrei Margeloiu, Nikola Simidjievski, Pietro Liò, Mateja JamnikAAAI 2023 · 19 citations
- InterpreTabNet: Distilling Predictive Signals from Tabular Data by Salient Feature InterpretationJacob Yoke Hong Si, Wendy Yusi Cheng, Michael Cooper, Rahul G. KrishnanICML 2024 · 15 citations
- Interpretable Prediction and Feature Selection for Survival AnalysisMike Van Ness, Madeleine UdellKDD 2025
- Self-Supervision Enhanced Feature Selection with Correlated GatesChanghee Lee, Fergus Imrie, Mihaela van der SchaarICLR 2022 · 26 citations
- Gradient-based Explanations for Deep Learning Survival ModelsSophie Hanna Langbein, Niklas Koenen, Marvin N. WrightICML 2025
