InterpreTabNet: Distilling Predictive Signals from Tabular Data by Salient Feature Interpretation
Jacob Yoke Hong Si, Wendy Yusi Cheng, Michael Cooper, Rahul G. Krishnan
摘要
Tabular data are omnipresent in various sectors of industries. Neural networks for tabular data such as TabNet have been proposed to make predictions while leveraging the attention mechanism for interpretability. However, the inferred attention masks are often dense, making it challenging to come up with rationales about the predictive signal. To remedy this, we propose InterpreTab-Net, a variant of the TabNet model that models the attention mechanism as a latent variable sampled from a Gumbel-Softmax distribution. This enables us to regularize the model to learn distinct concepts in the attention masks via a KL Divergence regularizer. It prevents overlapping feature selection by promoting sparsity which maximizes the model's efficacy and improves interpretability to determine the important features when predicting the outcome. To assist in the interpretation of feature interdependencies from our model, we employ a large language model (GPT-4) and use prompt engineering to map from the learned feature mask onto natural language text describing the learned signal. Through comprehensive experiments on real-world datasets, we demonstrate that InterpreTabNet outperforms previous methods for interpreting tabular data while attaining competitive accuracy.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Variational Uncertainty Decomposition for In-Context LearningI. Shavindra Jayasekera, Jacob Si, Filippo Valdettaro, Wenlong Chen 等NeurIPS 2025 · 被引用 7 次
- FEAT-KD: Learning Concise Representations for Single and Multi-Target Regression via TabNet Knowledge DistillationKei Sen Fong, Mehul MotaniICML 2025
它引用的顶会 Paper3
- TabNet: Attentive Interpretable Tabular LearningSercan Ö. Arik, Tomas PfisterAAAI 2021 · 被引用 2,148 次
- SubTab: Subsetting Features of Tabular Data for Self-Supervised Representation LearningTalip Ucar, Ehsan Hajiramezanali, Lindsay EdwardsNeurIPS 2021 · 被引用 189 次
- Net-DNF: Effective Deep Modeling of Tabular DataLiran Katzir, Gal Elidan, Ran El-YanivICLR 2021 · 被引用 40 次
相关 Paper
- MultiTab: A Scalable Foundation for Multitask Learning on Tabular DataDimitrios Sinodinos, Jack Yi Wei, Narges ArmanfardAAAI 2026 · 被引用 2 次
- TANGOS: Regularizing Tabular Neural Networks through Gradient Orthogonalization and SpecializationAlan Jeffares, Tennison Liu, Jonathan Crabbé, Fergus Imrie 等ICLR 2023 · 被引用 3 次
- Locally Sparse Neural Networks for Tabular Biomedical DataJunchen Yang, Ofir Lindenbaum, Yuval KlugerICML 2022 · 被引用 45 次
- Making Pre-trained Language Models Great on Tabular PredictionJiahuan Yan, Bo Zheng, Hongxia Xu, Yiheng Zhu 等ICLR 2024 · 被引用 72 次
- ARM-Net: Adaptive Relation Modeling Network for Structured DataShaofeng Cai, Kaiping Zheng, Gang Chen, H. V. Jagadish 等SIGMOD 2021 · 被引用 38 次
