DeepSeer: Interactive RNN Explanation and Debugging via State Abstraction
Zhijie Wang, Yuheng Huang, Da Song, Lei Ma, Tianyi Zhang
Abstract
Recurrent Neural Networks (RNNs) have been widely used in Natural Language Processing (NLP) tasks given its superior performance on processing sequential data. However, it is challenging to interpret and debug RNNs due to the inherent complexity and the lack of transparency of RNNs. While many explainable AI (XAI) techniques have been proposed for RNNs, most of them only support local explanations rather than global explanations. In this paper, we present DeepSeer, an interactive system that provides both global and local explanations of RNN behavior in multiple tightly-coordinated views for model understanding and debugging. The core of DeepSeer is a state abstraction method that bundles semantically similar hidden states in an RNN model and abstracts the model as a finite state machine. Users can explore the global model behavior by inspecting text patterns associated with each state and the transitions between states. Users can also dive into individual predictions by inspecting the state trace and intermediate prediction results of a given input. A between-subjects user study with 28 participants shows that, compared with a popular XAI technique, LIME, participants using DeepSeer made a deeper and more comprehensive assessment of RNN model behavior, identified the root causes of incorrect predictions more accurately, and came up with more actionable plans to improve the model performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9e0809fe-be78-4b96-b2f6-98e95e6fc856Cited by top-tier papers2
- EXMOS: Explanatory Model Steering through Multifaceted Explanations and Data ConfigurationsAditya Bhattacharya, Simone Stumpf, Lucija Gosak, Gregor Stiglic et al.CHI 2024 · 38 citations
- ReX: A Framework for Incorporating Temporal Information in Model-Agnostic Local Explanation TechniquesJunhao Liu, Xin ZhangAAAI 2025 · 6 citations
Builds on7
- Questioning the AI: Informing Design Practices for Explainable AI User ExperiencesQ. Vera Liao, Daniel M. Gruen, Sarah MillerCHI 2020 · 758 citations
- Interpreting Interpretability: Understanding Data Scientists' Use of Interpretability Tools for Machine LearningHarmanpreet Kaur, Harsha Nori, Samuel Jenkins, Rich Caruana et al.CHI 2020 · 541 citations
- Debugging Tests for Model ExplanationsJulius Adebayo, Michael Muelly, Ilaria Liccardi, Been KimNeurIPS 2020 · 209 citations
- ExplAIn Yourself! Transparency for Positive UX in Autonomous DrivingTobias Schneider, Joana Hois, Alischa Rosenstein, Sabiha Ghellal et al.CHI 2021 · 74 citations
- Weighted Automata Extraction from Recurrent Neural Networks via Regression on State SpacesTakamasa Okudono, Masaki Waga, Taro Sekiyama, Ichiro HasuoAAAI 2020 · 44 citations
Related papers
- AdaAX: Explaining Recurrent Neural Networks by Learning Automata with Adaptive StatesDat Hong, Alberto Maria Segre, Tong WangKDD 2022 · 3 citations
- Decision-Guided Weighted Automata Extraction from Recurrent Neural NetworksXiyue Zhang, Xiaoning Du, Xiaofei Xie, Lei Ma et al.AAAI 2021 · 25 citations
- VIME: Visual Interactive Model Explorer for Identifying Capabilities and Limitations of Machine Learning Models for Sequential Decision-MakingAnindya Das Antar, Somayeh Molaei, Yan-Ying Chen, Matthew L. Lee et al.UIST 2024 · 3 citations
- How Can I Explain This to You? An Empirical Study of Deep Neural Network Explanation MethodsJeya Vikranth Jeyakumar, Joseph Noor, Yu-Hsi Cheng, Luis Garcia et al.NeurIPS 2020 · 173 citations
- Learning Explainable Linguistic Expressions with Neural Inductive Logic Programming for Sentence ClassificationPrithviraj Sen, Marina Danilevsky, Yunyao Li, Siddhartha Brahma et al.EMNLP 2020 · 12 citations
