Explainable Neural Networks with Guarantee: A Sparse Estimation Approach
Antoine Ledent, Peng Liu
摘要
Balancing predictive power and interpretability has long been a challenging research area, particularly in powerful yet complex models like neural networks, where nonlinearity obstructs direct interpretation. This paper introduces a novel approach to constructing an explainable neural network that harmonizes predictiveness and explainability. Our model, termed SparXnet, is designed as a linear combination of a sparse set of jointly learned features, each derived from a different trainable function applied to a single 1-dimensional input feature. Leveraging the ability to learn arbitrarily complex relationships, our neural network architecture enables automatic selection of a sparse set of important features, with the final prediction being a linear combination of rescaled versions of these features. We demonstrate the ability to select significant features while maintaining comparable predictive performance and direct interpretability through extensive experiments on synthetic and real-world datasets. We also provide theoretical analysis on the generalization bounds of our framework, which is favorably linear in the number of selected features and only logarithmic in the number of input features. We further lift any dependence of sample complexity on the number of parameters or the architectural details under very mild conditions. Our research paves the way for further research on sparse and explainable neural networks with guarantee.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper8
- Neural Additive Models: Interpretable Machine Learning with Neural NetsRishabh Agarwal, Levi Melnick, Nicholas Frosst, Xuezhou Zhang 等NeurIPS 2021 · 被引用 663 次
- Norm-Based Generalisation Bounds for Deep Multi-Class Convolutional Neural NetworksAntoine Ledent, Waleed Mustafa, Yunwen Lei, Marius KloftAAAI 2021 · 被引用 24 次
- On Measuring Excess Capacity in Neural NetworksFlorian Graf, Sebastian Zeng, Bastian Rieck, Marc Niethammer 等NeurIPS 2022 · 被引用 13 次
- Exploring and Adapting Chinese GPT to Pinyin Input MethodMinghuan Tan, Yong Dai, Duyu Tang, Zhangyin Feng 等ACL 2022 · 被引用 13 次
- Beyond Smoothness: Incorporating Low-Rank Analysis into Nonparametric Density EstimationRobert A. Vandermeulen, Antoine LedentNeurIPS 2021 · 被引用 12 次
相关 Paper
- The Contextual Lasso: Sparse Linear Models via Deep Neural NetworksRyan Thompson, Amir Dezfouli, Robert KohnNeurIPS 2023 · 被引用 8 次
- Kernel Logistic Regression Approximation of an Understandable ReLU Neural NetworkMarie Guyomard, Susana Barbosa, Lionel FillatreICML 2023
- NIMO: a Nonlinear Interpretable MOdelShijian Xu, Marcello Massimo Negri, Volker RothICLR 2026 · 被引用 1 次
- Learning Accurate and Interpretable Decision Rule Sets from Neural NetworksLitao Qiao, Weijia Wang, Bill LinAAAI 2021 · 被引用 53 次
- Leveraging Sparse Linear Layers for Debuggable Deep NetworksEric Wong, Shibani Santurkar, Aleksander MadryICML 2021 · 被引用 101 次
