Human-in-the-loop Extraction of Interpretable Concepts in Deep Learning Models
Zhenge Zhao, Panpan Xu, Carlos Scheidegger, Liu Ren
Abstract
The interpretation of deep neural networks (DNNs) has become a key topic as more and more people apply them to solve various problems and making critical decisions. Concept-based explanations have recently become a popular approach for post-hoc interpretation of DNNs. However, identifying human-understandable visual concepts that affect model decisions is a challenging task that is not easily addressed with automatic approaches. We present a novel human-in-the-Ioop approach to generate user-defined concepts for model interpretation and diagnostics. Central to our proposal is the use of active learning, where human knowledge and feedback are combined to train a concept extractor with very little human labeling effort. We integrate this process into an interactive system, ConceptExtract. Through two case studies, we show how our approach helps analyze model behavior and extract human-friendly concepts for different machine learning tasks and datasets and how to use these concepts to understand the predictions, compare model performance and make suggestions for model refinement. Quantitative experiments show that our active learning approach can accurately extract meaningful visual concepts. More importantly, by identifying visual concepts that negatively affect model performance, we develop the corresponding data augmentation strategy that consistently improves model performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ebf0f618-21f0-4495-b768-0e441ca67123Cited by top-tier papers8
- ConceptExplainer: Interactive Explanation for Deep Neural Networks from a Concept PerspectiveJinbin Huang, Aditi Mishra, Bum Chul Kwon, Chris BryanIEEE VIS 2022 · 46 citations
- DendroMap: Visual Exploration of Large-Scale Image Datasets for Machine Learning with TreemapsDonald Bertucci, Md Montaser Hamid, Yashwanthi Anand, Anita Ruangrotsakun et al.IEEE VIS 2022 · 34 citations
- LLM Comparator: Interactive Analysis of Side-by-Side Evaluation of Large Language ModelsMinsuk Kahng, Ian Tenney, Mahima Pushkarna, Michael Xieyang Liu et al.IEEE VIS 2024 · 23 citations
- VLSlice: Interactive Vision-and-Language Slice DiscoveryEric Slyman, Minsuk Kahng, Stefan LeeICCV 2023 · 11 citations
- ESCAPE: Countering Systematic Errors from Machine's Blind Spots via Interactive Visual AnalysisYongsu Ahn, Yu-Ru Lin, Panpan Xu, Zeng DaiCHI 2023 · 10 citations
Related papers
- Explanation-based Data Augmentation for Image ClassificationSandareka Wickramanayake, Wynne Hsu, Mong-Li LeeNeurIPS 2021 · 24 citations
- Overlooked Factors in Concept-Based Explanations: Dataset Choice, Concept Learnability, and Human CapabilityVikram V. Ramaswamy, Sunnie S. Y. Kim, Ruth Fong, Olga RussakovskyCVPR 2023
- How can Explainability Methods be Used to Support Bug Identification in Computer Vision Models?Agathe Balayn, Natasa Rikalo, Christoph Lofi, Jie Yang et al.CHI 2022 · 22 citations
- A Peek Into the Reasoning of Neural Networks: Interpreting With Structural Visual ConceptsYunhao Ge, Yao Xiao, Zhi Xu, Meng Zheng et al.CVPR 2021
- Towards Trustable Skin Cancer Diagnosis via Rewriting Model's DecisionSiyuan Yan, Zhen Yu, Xuelin Zhang, Dwarikanath Mahapatra et al.CVPR 2023
