FEAT-KD: Learning Concise Representations for Single and Multi-Target Regression via TabNet Knowledge Distillation
Kei Sen Fong, Mehul Motani
Abstract
In this work, we propose a novel approach that combines the strengths of FEAT and Tab-Net through knowledge distillation (KD), which we term FEAT-KD. FEAT is an intrinsically interpretable machine learning (ML) algorithm that constructs a weighted linear combination of concisely-represented features discovered via genetic programming optimization, which can often be inefficient. FEAT-KD leverages TabNet's deeplearning-based optimization and feature selection mechanisms instead. FEAT-KD finds a weighted linear combination of concisely-represented, symbolic features that are derived from piece-wise distillation of a trained TabNet model. We analyze FEAT-KD on regression tasks from two perspectives: (i) compared to TabNet, FEAT-KD significantly reduces model complexity while retaining competitive predictive performance, effectively converting a black-box deep learning model into a more interpretable white-box representation, (ii) compared to FEAT, our method consistently outperforms in prediction accuracy, produces more compact models, and reduces the complexity of learned symbolic expressions. In addition, we demonstrate that FEAT-KD easily supports multitarget regression, in which the shared features contribute to the interpretability of the system. Our results suggest that FEAT-KD is a promising direction for interpretable ML, bridging the gap between deep learning's predictive power and the intrinsic transparency of symbolic models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0cce479f-2c2c-4e17-ab7a-14c41c551716Builds on7
- TabNet: Attentive Interpretable Tabular LearningSercan Ö. Arik, Tomas PfisterAAAI 2021 · 2,148 citations
- Discovering Symbolic Models from Deep Learning with Inductive BiasesMiles D. Cranmer, Alvaro Sanchez-Gonzalez, Peter W. Battaglia, Rui Xu et al.NeurIPS 2020 · 736 citations
- Deep symbolic regression: Recovering mathematical expressions from data via risk-seeking policy gradientsBrenden K. Petersen, Mikel Landajuela, T. Nathan Mundhenk, Cláudio Prata Santiago et al.ICLR 2021 · 444 citations
- Neural Symbolic Regression that scalesLuca Biggio, Tommaso Bendinelli, Alexander Neitz, Aurélien Lucchi et al.ICML 2021 · 251 citations
- COGAM: Measuring and Moderating Cognitive Load in Machine Learning Model ExplanationsAshraf M. Abdul, Christian von der Weth, Mohan S. Kankanhalli, Brian Y. LimCHI 2020 · 92 citations
Related papers
- Symbolic Metamodels for Interpreting Black-Boxes Using Primitive FunctionsMahed Abroshan, Saumitra Mishra, Mohammad Mahdi KhaliliAAAI 2023 · 5 citations
- On the Impact of Knowledge Distillation for Model InterpretabilityHyeongrok Han, Siwon Kim, Hyun-Soo Choi, Sungroh YoonICML 2023 · 13 citations
- Synergizing Large Language Models and Knowledge-Based Reasoning for Interpretable Feature EngineeringMohamed Bouadi, Arta Alavi, Salima Benbernou, Mourad OuziriWWW 2025 · 2 citations
- InterpreTabNet: Distilling Predictive Signals from Tabular Data by Salient Feature InterpretationJacob Yoke Hong Si, Wendy Yusi Cheng, Michael Cooper, Rahul G. KrishnanICML 2024 · 15 citations
- An Interpretable Approach to the Solutions of High-Dimensional Partial Differential EquationsLulu Cao, Yufei Liu, Zhenzhong Wang, Dejun Xu et al.AAAI 2024 · 15 citations
