Efficient Understanding of Machine Learning Model Mispredictions
Martin Eberlein, Jürgen Cito, Lars Grunske
Abstract
Mispredictions by machine learning components can have severe consequences, especially in safety-critical and mission-critical software systems. Therefore, understanding and debugging these mispredictions is a crucial part of the development process for systems that use machine learning components. Previous research has successfully applied methods that identify when a model's predictions may be unreliable by generating a rule set that links feature values to prediction errors. However, current state-of-the-art rule set approaches require significant computational resources, particularly for large data sets. To address these high computational demands, we propose a strategy to identify and focus only on the most influential features that lead to mispredictions. Additionally, to improve the accuracy of mispredictions diagnosis, we replace traditional rule-based approaches with decision tree learning. We evaluate our tool MMDFAST across 11 diverse real-world data sets. The results show that focusing on influential features with decision trees improves the accuracy of misprediction explanations, while significantly reducing computational demands in all scenarios. Thus, MMDFAST produces better results much faster, making it more efficient and effective for generating misprediction explanations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 909b4e4d-a13e-457c-8d59-95066151dadbBuilds on5
- Regression Fuzzing for Deep Learning SystemsHanmo You, Zan Wang, Junjie Chen, Shuang Liu et al.ICSE 2023 · 28 citations
- Explaining mispredictions of machine learning models using rule inductionJürgen Cito, Isil Dillig, Seohyun Kim, Vijayaraghavan Murali et al.FSE 2021 · 26 citations
- DENAS: automated rule generation by knowledge extraction from neural networksSimin Chen, Soroush Bateni, Sampath Grandhi, Xiaodi Li et al.FSE 2020 · 22 citations
- Semantic DebuggingMartin Eberlein, Marius Smytzek, Dominic Steinhöfel, Lars Grunske et al.FSE 2023 · 13 citations
- Leveraging Feature Bias for Scalable Misprediction Explanation of Machine Learning ModelsJiri Gesi, Xinyun Shen, Yunfan Geng, Qihong Chen et al.ICSE 2023 · 8 citations
Related papers
- A Scalable Two Stage Approach to Computing Optimal Decision SetsAlexey Ignatiev, Edward Lam, Peter J. Stuckey, João Marques-SilvaAAAI 2021 · 17 citations
- Selective ExplanationsLucas Monteiro Paes, Dennis Wei, Flávio P. CalmonNeurIPS 2024 · 4 citations
- Computing Rule-Based Explanations by Leveraging CounterfactualsZixuan Geng, Maximilian Schleich, Dan SuciuVLDB 2023 · 7 citations
- "Why is 'Chicago' deceptive?" Towards Building Model-Driven Tutorials for HumansVivian Lai, Han Liu, Chenhao TanCHI 2020 · 113 citations
- DECE: Decision Explorer with Counterfactual Explanations for Machine Learning ModelsFurui Cheng, Yao Ming, Huamin QuIEEE VIS 2020 · 118 citations
