Symbolic Metamodels for Interpreting Black-Boxes Using Primitive Functions
Mahed Abroshan, Saumitra Mishra, Mohammad Mahdi Khalili
Abstract
One approach for interpreting black-box machine learning models is to find a global approximation of the model using simple interpretable functions, which is called a metamodel (a model of the model). Approximating the black-box with a metamodel can be used to 1) estimate instance-wise feature importance; 2) understand the functional form of the model; 3) analyze feature interactions. In this work, we propose a new method for finding interpretable metamodels. Our approach utilizes Kolmogorov superposition theorem, which expresses multivariate functions as a composition of univariate functions (our primitive parameterized functions). This composition can be represented in the form of a tree. Inspired by symbolic regression, we use a modified form of genetic programming to search over different tree configurations. Gradient descent (GD) is used to optimize the parameters of a given configuration. Our method is a novel memetic algorithm that uses GD not only for training numerical constants but also for the training of building blocks. Using several experiments, we show that our method outperforms recent metamodeling approaches suggested for interpreting black-boxes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 06269ef9-3e25-40ee-9dd0-3548b80cf330Builds on2
- Deep symbolic regression: Recovering mathematical expressions from data via risk-seeking policy gradientsBrenden K. Petersen, Mikel Landajuela, T. Nathan Mundhenk, Cláudio Prata Santiago et al.ICLR 2021 · 444 citations
- Learning outside the Black-Box: The pursuit of interpretable modelsJonathan Crabbé, Yao Zhang, William R. Zame, Mihaela van der SchaarNeurIPS 2020 · 30 citations
Related papers
- FEAT-KD: Learning Concise Representations for Single and Multi-Target Regression via TabNet Knowledge DistillationKei Sen Fong, Mehul MotaniICML 2025
- Improving Memory Efficiency for Training KANs via Meta LearningZhangchi Zhao, Jun Shu, Deyu Meng, Zongben XuICML 2025
- Prediction via Shapley Value RegressionAmr Alkhatib, Roman Bresson, Henrik Boström, Michalis VazirgiannisICML 2025
- Symbolic Regression via Deep Reinforcement Learning Enhanced Genetic Programming SeedingT. Nathan Mundhenk, Mikel Landajuela, Ruben Glatt, Cláudio P. Santiago et al.NeurIPS 2021 · 95 citations
- SOInter: A Novel Deep Energy-Based Interpretation Method for Explaining Structured Output ModelsSeyyede Fatemeh Seyyedsalehi, Mahdieh Soleymani Baghshah, Hamid R. RabieeICLR 2024
