Neural Additive Models: Interpretable Machine Learning with Neural Nets
Rishabh Agarwal, Levi Melnick, Nicholas Frosst, Xuezhou Zhang, Benjamin J. Lengerich, Rich Caruana, Geoffrey E. Hinton
Abstract
Deep neural networks (DNNs) are powerful black-box predictors that have achieved impressive performance on a wide variety of tasks. However, their accuracy comes at the cost of intelligibility: it is usually unclear how they make their decisions. This hinders their applicability to high stakes decision-making domains such as healthcare. We propose Neural Additive Models (NAMs) which combine some of the expressivity of DNNs with the inherent intelligibility of generalized additive models. NAMs learn a linear combination of neural networks that each attend to a single input feature. These networks are trained jointly and can learn arbitrarily complex relationships between their input feature and the output. Our experiments on regression and classification datasets show that NAMs are more accurate than widely used intelligible models such as logistic regression and shallow decision trees. They perform similarly to existing state-of-the-art generalized additive models in accuracy, but are more flexible because they are based on neural nets instead of boosted trees. To demonstrate this, we show how NAMs can be used for multitask learning on synthetic data and on the COMPAS recidivism data due to their composability, and demonstrate that the differentiability of NAMs allows them to train more complex interpretable models for COVID-19. Source code is available at neural-additive-models.github.io.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext af9cbd23-a95e-442c-b704-a731752812bbCited by top-tier papers77
- Separable Physics-Informed Neural NetworksJunwoo Cho, Seungtae Nam, Hyunmo Yang, Seok-Bae Yun et al.NeurIPS 2023 · 138 citations
- Additive MIL: Intrinsically Interpretable Multiple Instance Learning for PathologySyed Ashar Javed, Dinkar Juyal, Harshith Padigela, Amaro Taylor-Weiner et al.NeurIPS 2022 · 124 citations
- NODE-GAM: Neural Generalized Additive Model for Interpretable Deep LearningChun-Hao Chang, Rich Caruana, Anna GoldenbergICLR 2022 · 114 citations
- Neural Basis Models for InterpretabilityFilip Radenovic, Abhimanyu Dubey, Dhruv MahajanNeurIPS 2022 · 82 citations
- Encoding Time-Series Explanations through Self-Supervised Model Behavior ConsistencyOwen Queen, Tom Hartvigsen, Teddy Koker, Huan He et al.NeurIPS 2023 · 55 citations
Builds on1
Related papers
- The Intelligible and Effective Graph Neural Additive NetworkMaya Bechler-Speicher, Amir Globerson, Ran Gilad-BachrachNeurIPS 2024 · 31 citations
- Gaussian Process Neural Additive ModelsWei Zhang, Brian Barr, John PaisleyAAAI 2024 · 16 citations
- CAT: Interpretable Concept-based Taylor Additive ModelsViet Duong, Qiong Wu, Zhengyi Zhou, Hongjue Zhao et al.KDD 2024 · 6 citations
- Scalable Interpretability via PolynomialsAbhimanyu Dubey, Filip Radenovic, Dhruv MahajanNeurIPS 2022 · 42 citations
- Curve Your Enthusiasm: Concurvity Regularization in Differentiable Generalized Additive ModelsJulien Siems, Konstantin Ditschuneit, Winfried Ripken, Alma Lindborg et al.NeurIPS 2023 · 15 citations
