Listen to Interpret: Post-hoc Interpretability for Audio Networks with NMF
Jayneel Parekh, Sanjeel Parekh, Pavlo Mozharovskyi, Florence d'Alché-Buc, Gaël Richard
Abstract
This paper tackles post-hoc interpretability for audio processing networks. Our goal is to interpret decisions of a network in terms of high-level audio objects that are also listenable for the end-user. To this end, we propose a novel interpreter design that incorporates non-negative matrix factorization (NMF). In particular, a carefully regularized interpreter module is trained to take hidden layer representations of the targeted network as input and produce time activations of pre-learnt NMF components as intermediate outputs. Our methodology allows us to generate intuitive audio-based interpretations that explicitly enhance parts of the input signal most relevant for a network's decision. We demonstrate our method's applicability on popular benchmarks, including a real-world multi-label classification task.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6de3da02-113a-4cd3-87dc-2de2cfce7214Cited by top-tier papers7
- A Holistic Approach to Unifying Automatic Concept Extraction and Concept Importance EstimationThomas Fel, Victor Boutin, Louis Béthune, Rémi Cadène et al.NeurIPS 2023 · 125 citations
- Listenable Maps for Audio ClassifiersFrancesco Paissan, Mirco Ravanelli, Cem SubakanICML 2024 · 13 citations
- Listenable Maps for Zero-Shot Audio ClassifiersFrancesco Paissan, Luca Della Libera, Mirco Ravanelli, Cem SubakanNeurIPS 2024 · 5 citations
- TimeSAE: Causal Sparse Decoding for Faithful Explanations of Black-Box Time Series ModelsKhalid Oublal, Quentin Bouniot, Qi Gan, Stephan Clemencon et al.ICML 2026 · 2 citations
- Enhancing Uncertainty Estimation and Interpretability with Bayesian Non-negative Decision LayerXinyue Hu, Zhibin Duan, Bo Chen, Mingyuan ZhouICLR 2025
Builds on4
- Restricting the Flow: Information Bottlenecks for AttributionKarl Schulz, Leon Sixt, Federico Tombari, Tim LandgrafICLR 2020 · 220 citations
- What went wrong and when? Instance-wise feature importance for time-series black-box modelsSana Tonekaboni, Shalmali Joshi, Kieran Campbell, David Duvenaud et al.NeurIPS 2020 · 94 citations
- A Framework to Learn with InterpretationJayneel Parekh, Pavlo Mozharovskyi, Florence d'Alché-BucNeurIPS 2021 · 35 citations
- Explaining A Black-box By Using A Deep Variational Information Bottleneck ApproachSeo-Jin Bang, Pengtao Xie, Heewook Lee, Wei Wu et al.AAAI 2021 · 33 citations
Related papers
- FACE: Faithful Automatic Concept ExtractionDipkamal Bhusal, Michael Clifford, Sara Rampazzi, Nidhi RastogiNeurIPS 2025 · 11 citations
- Non-negative Contrastive LearningYifei Wang, Qi Zhang, Yaoyu Guo, Yisen WangICLR 2024 · 18 citations
- AND: Audio Network Dissection for Interpreting Deep Acoustic ModelsTung-Yu Wu, Yu-Xiang Lin, Tsui-Wei WengICML 2024 · 6 citations
- Constructing Interpretable Features from Compositional Neuron GroupsOr David Shafran, Atticus Geiger, Mor GevaACL 2026 · 4 citations
- LEAF: A Learnable Frontend for Audio ClassificationNeil Zeghidour, Olivier Teboul, Félix de Chaumont Quitry, Marco TagliasacchiICLR 2021 · 181 citations
