Listen to Interpret: Post-hoc Interpretability for Audio Networks with NMF
Jayneel Parekh, Sanjeel Parekh, Pavlo Mozharovskyi, Florence d'Alché-Buc, Gaël Richard
摘要
This paper tackles post-hoc interpretability for audio processing networks. Our goal is to interpret decisions of a network in terms of high-level audio objects that are also listenable for the end-user. To this end, we propose a novel interpreter design that incorporates non-negative matrix factorization (NMF). In particular, a carefully regularized interpreter module is trained to take hidden layer representations of the targeted network as input and produce time activations of pre-learnt NMF components as intermediate outputs. Our methodology allows us to generate intuitive audio-based interpretations that explicitly enhance parts of the input signal most relevant for a network's decision. We demonstrate our method's applicability on popular benchmarks, including a real-world multi-label classification task.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- A Holistic Approach to Unifying Automatic Concept Extraction and Concept Importance EstimationThomas Fel, Victor Boutin, Louis Béthune, Rémi Cadène 等NeurIPS 2023 · 被引用 125 次
- Listenable Maps for Audio ClassifiersFrancesco Paissan, Mirco Ravanelli, Cem SubakanICML 2024 · 被引用 13 次
- Listenable Maps for Zero-Shot Audio ClassifiersFrancesco Paissan, Luca Della Libera, Mirco Ravanelli, Cem SubakanNeurIPS 2024 · 被引用 5 次
- TimeSAE: Causal Sparse Decoding for Faithful Explanations of Black-Box Time Series ModelsKhalid Oublal, Quentin Bouniot, Qi Gan, Stephan Clemencon 等ICML 2026 · 被引用 2 次
- Enhancing Uncertainty Estimation and Interpretability with Bayesian Non-negative Decision LayerXinyue Hu, Zhibin Duan, Bo Chen, Mingyuan ZhouICLR 2025
它引用的顶会 Paper4
- Restricting the Flow: Information Bottlenecks for AttributionKarl Schulz, Leon Sixt, Federico Tombari, Tim LandgrafICLR 2020 · 被引用 220 次
- What went wrong and when? Instance-wise feature importance for time-series black-box modelsSana Tonekaboni, Shalmali Joshi, Kieran Campbell, David Duvenaud 等NeurIPS 2020 · 被引用 94 次
- A Framework to Learn with InterpretationJayneel Parekh, Pavlo Mozharovskyi, Florence d'Alché-BucNeurIPS 2021 · 被引用 35 次
- Explaining A Black-box By Using A Deep Variational Information Bottleneck ApproachSeo-Jin Bang, Pengtao Xie, Heewook Lee, Wei Wu 等AAAI 2021 · 被引用 33 次
相关 Paper
- FACE: Faithful Automatic Concept ExtractionDipkamal Bhusal, Michael Clifford, Sara Rampazzi, Nidhi RastogiNeurIPS 2025 · 被引用 11 次
- Non-negative Contrastive LearningYifei Wang, Qi Zhang, Yaoyu Guo, Yisen WangICLR 2024 · 被引用 18 次
- AND: Audio Network Dissection for Interpreting Deep Acoustic ModelsTung-Yu Wu, Yu-Xiang Lin, Tsui-Wei WengICML 2024 · 被引用 6 次
- Constructing Interpretable Features from Compositional Neuron GroupsOr David Shafran, Atticus Geiger, Mor GevaACL 2026 · 被引用 4 次
- LEAF: A Learnable Frontend for Audio ClassificationNeil Zeghidour, Olivier Teboul, Félix de Chaumont Quitry, Marco TagliasacchiICLR 2021 · 被引用 181 次
