NIMO: a Nonlinear Interpretable MOdel
Shijian Xu, Marcello Massimo Negri, Volker Roth
Abstract
Deep learning has achieved remarkable success across many domains, but it has also created a growing demand for interpretability in model predictions. Although many explainable machine learning methods have been proposed, post-hoc explanations lack guaranteed fidelity and are sensitive to hyperparameter choices, highlighting the appeal of inherently interpretable models. For example, linear regression provides clear feature effects through its coefficients. However, such models are often outperformed by more complex neural networks (NNs) that usually lack inherent interpretability. To address this dilemma, we introduce NIMO, a framework that combines inherent interpretability with the expressive power of neural networks. Building on the simple linear regression, NIMO is able to provide flexible and intelligible feature effects. Relevantly, we develop an optimization method based on parameter elimination, that allows for optimizing the NN parameters and linear coefficients effectively and efficiently. By relying on adaptive ridge regression we can easily incorporate sparsity as well. We show empirically that our model can provide faithful and intelligible feature effects while maintaining good predictive performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6ce890d0-1c2d-4559-a14d-2dd1b60ce92eBuilds on6
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell et al.NeurIPS 2020 · 4,008 citations
- Neural Additive Models: Interpretable Machine Learning with Neural NetsRishabh Agarwal, Levi Melnick, Nicholas Frosst, Xuezhou Zhang et al.NeurIPS 2021 · 663 citations
- NODE-GAM: Neural Generalized Additive Model for Interpretable Deep LearningChun-Hao Chang, Rich Caruana, Anna GoldenbergICLR 2022 · 114 citations
- The Contextual Lasso: Sparse Linear Models via Deep Neural NetworksRyan Thompson, Amir Dezfouli, Robert KohnNeurIPS 2023 · 8 citations
- Interpretable Mesomorphic Networks for Tabular DataArlind Kadra, Sebastian Pineda-Arango, Josif GrabockaNeurIPS 2024 · 6 citations
Related papers
- Leveraging Sparse Linear Layers for Debuggable Deep NetworksEric Wong, Shibani Santurkar, Aleksander MadryICML 2021 · 101 citations
- Explainable Neural Networks with Guarantee: A Sparse Estimation ApproachAntoine Ledent, Peng LiuAAAI 2025 · 1 citation
- Regularizing Black-box Models for Improved InterpretabilityGregory Plumb, Maruan Al-Shedivat, Ángel Alexander Cabrera, Adam Perer et al.NeurIPS 2020 · 90 citations
- Neural Basis Models for InterpretabilityFilip Radenovic, Abhimanyu Dubey, Dhruv MahajanNeurIPS 2022 · 82 citations
- Gnothi Seauton: Empowering Faithful Self-Interpretability in Black-Box TransformersShaobo Wang, Hongxuan Tang, Mingyang Wang, Hongrui Zhang et al.ICLR 2025
