Interpreting Internal Activation Patterns in Deep Temporal Neural Networks by Finding Prototypes
Sohee Cho, Wonjoon Chang, Ginkyeng Lee, Jaesik Choi
摘要
Deep neural networks have demonstrated competitive performance in classification tasks for sequential data. However, it remains difficult to understand which temporal patterns the internal channels of deep neural networks capture for decision-making in sequential data. To address this issue, we propose a new framework with which to visualize temporal representations learned in deep neural networks without hand-crafted segmentation labels. Given input data, our framework extracts highly activated temporal regions that contribute to activating internal nodes and characterizes such regions by prototype selection method based on Maximum Mean Discrepancy. Representative temporal patterns referred to here as Prototypes of Temporally Activated Patterns (PTAP) provide core examples of subsequences in the sequential data for interpretability. We also analyze the role of each channel by Value-LRP plots using representative prototypes and the distribution of the input attribution. Input attribution plots give visual information to recognize the shapes focused on by the channel for decision-making.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper4
- MIX: A Multi-view Time-Frequency Interactive Explanation Framework for Time Series ClassificationViet-Hung Tran, Ngoc Phu Doan, Zichi Zhang, Tuan Dung Pham 等NeurIPS 2025 · 被引用 3 次
- Understanding Distributed Representations of Concepts in Deep Neural Networks without SupervisionWonjoon Chang, Dahee Kwon, Jaesik ChoiAAAI 2024 · 被引用 2 次
- IIEU: Rethinking Neural Feature Activation from Decision-MakingSudong CaiICCV 2023 · 被引用 1 次
- Unified Time Series Explanations via Amortized Optimization and Instance-level Multi-Expert Knowledge DistillationViet-Hung Tran, Zichi Zhang, Ngoc Doan, Xuan Hoang Nguyen 等ICML 2026
相关 Paper
- Interpretable Image Classification via Non-parametric Part Prototype LearningZhijie Zhu, Lei Fan, Maurice Pagnucco, Yang SongCVPR 2025
- Relative Attributing Propagation: Interpreting the Comparative Contributions of Individual Units in Deep Neural NetworksWoo-Jeoung Nam, Shir Gur, Jaesik Choi, Lior Wolf 等AAAI 2020 · 被引用 109 次
- ProtoTS: Learning Hierarchical Prototypes for Explainable Time Series ForecastingZiheng Peng, Shijie Ren, Xinyue Gu, Linxiao Yang 等ICLR 2026 · 被引用 2 次
- Learning to Select Prototypical Parts for Interpretable Sequential Data ModelingYifei Zhang, Neng Gao, Cunqing MaAAAI 2023 · 被引用 10 次
- SAEs-BrainMap: Unveiling the Emergence of Specialized Concepts in Deep Models via Brain AlignmentZiming Mao, Jia Xu, Wenxuan Pan, Mufan Xue 等ICML 2026
