Efficient Understanding of Machine Learning Model Mispredictions
Martin Eberlein, Jürgen Cito, Lars Grunske
摘要
Mispredictions by machine learning components can have severe consequences, especially in safety-critical and mission-critical software systems. Therefore, understanding and debugging these mispredictions is a crucial part of the development process for systems that use machine learning components. Previous research has successfully applied methods that identify when a model's predictions may be unreliable by generating a rule set that links feature values to prediction errors. However, current state-of-the-art rule set approaches require significant computational resources, particularly for large data sets. To address these high computational demands, we propose a strategy to identify and focus only on the most influential features that lead to mispredictions. Additionally, to improve the accuracy of mispredictions diagnosis, we replace traditional rule-based approaches with decision tree learning. We evaluate our tool MMDFAST across 11 diverse real-world data sets. The results show that focusing on influential features with decision trees improves the accuracy of misprediction explanations, while significantly reducing computational demands in all scenarios. Thus, MMDFAST produces better results much faster, making it more efficient and effective for generating misprediction explanations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Regression Fuzzing for Deep Learning SystemsHanmo You, Zan Wang, Junjie Chen, Shuang Liu 等ICSE 2023 · 被引用 28 次
- Explaining mispredictions of machine learning models using rule inductionJürgen Cito, Isil Dillig, Seohyun Kim, Vijayaraghavan Murali 等FSE 2021 · 被引用 26 次
- DENAS: automated rule generation by knowledge extraction from neural networksSimin Chen, Soroush Bateni, Sampath Grandhi, Xiaodi Li 等FSE 2020 · 被引用 22 次
- Semantic DebuggingMartin Eberlein, Marius Smytzek, Dominic Steinhöfel, Lars Grunske 等FSE 2023 · 被引用 13 次
- Leveraging Feature Bias for Scalable Misprediction Explanation of Machine Learning ModelsJiri Gesi, Xinyun Shen, Yunfan Geng, Qihong Chen 等ICSE 2023 · 被引用 8 次
相关 Paper
- A Scalable Two Stage Approach to Computing Optimal Decision SetsAlexey Ignatiev, Edward Lam, Peter J. Stuckey, João Marques-SilvaAAAI 2021 · 被引用 17 次
- Selective ExplanationsLucas Monteiro Paes, Dennis Wei, Flávio P. CalmonNeurIPS 2024 · 被引用 4 次
- Computing Rule-Based Explanations by Leveraging CounterfactualsZixuan Geng, Maximilian Schleich, Dan SuciuVLDB 2023 · 被引用 7 次
- "Why is 'Chicago' deceptive?" Towards Building Model-Driven Tutorials for HumansVivian Lai, Han Liu, Chenhao TanCHI 2020 · 被引用 113 次
- DECE: Decision Explorer with Counterfactual Explanations for Machine Learning ModelsFurui Cheng, Yao Ming, Huamin QuIEEE VIS 2020 · 被引用 118 次
