Label-Descriptive Patterns and Their Application to Characterizing Classification Errors
Michael A. Hedderich, Jonas Fischer, Dietrich Klakow, Jilles Vreeken
Abstract
State-of-the-art deep learning methods achieve human-like performance on many tasks, but make errors nevertheless. Characterizing these errors in easily interpretable terms gives insight into whether a model is prone to making systematic errors, but also gives a way to act and improve the model. In this paper we propose a method that allows us to do so for arbitrary classifiers by mining a small set of patterns that together succinctly describe the input data that is partitioned according to correctness of prediction. We show this is an instance of the more general label description problem, which we formulate in terms of the Minimum Description Length principle. To discover good pattern sets we propose the efficient and hyperparameter-free Premise algorithm, which through an extensive set of experiments we show on both synthetic and real-world data performs very well in practice; unlike existing solutions it ably recovers ground truth patterns, even on highly imbalanced data over many unique items, or where patterns are only weakly associated to labels. Through two real-world case studies we confirm that Premise gives clear and actionable insight into the systematic errors made by modern NLP classifiers.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 47a98df8-7da9-4a88-9473-e009a77a5bdcCited by top-tier papers6
- Finding Interpretable Class-Specific Patterns through Efficient Neural SearchNils Philipp Walter, Jonas Fischer, Jilles VreekenAAAI 2024 · 8 citations
- Evaluating Robustness of Large Language Models Against Multilingual Typographical ErrorsRaoyuan Zhao, Yihong Liu, Lena Altinger, Hinrich Schütze et al.ACL 2026 · 6 citations
- Divisi: Interactive Search and Visualization for Scalable Exploratory Subgroup AnalysisVenkatesh Sivaraman, Zexuan Li, Adam PererCHI 2025 · 5 citations
- CohEx: A Generalized Framework for Cohort ExplanationFanyu Meng, Xin Liu, Zhaodan Kong, Xin ChenAAAI 2025 · 4 citations
- Intrinsic User-Centric Interpretability through Global Mixture of ExpertsVinitra Swamy, Syrielle Montariol, Julian Blackwell, Jibril Frej et al.ICLR 2025
Builds on3
- Beyond Accuracy: Behavioral Testing of NLP Models with CheckListMarco Túlio Ribeiro, Tongshuang Wu, Carlos Guestrin, Sameer SinghACL 2020 · 51 citations
- What's in the Box? Exploring the Inner Life of Neural Networks with Robust RulesJonas Fischer, Anna Oláh, Jilles VreekenICML 2021 · 11 citations
- Discovering Succinct Pattern Sets Expressing Co-Occurrence and Mutual ExclusivityJonas Fischer, Jilles VreekenKDD 2020 · 6 citations
Related papers
- DISCERN: Decoding Systematic Errors in Natural Language for Text ClassifiersRakesh R. Menon, Shashank SrivastavaEMNLP 2024
- PRIME: Prioritizing Interpretability in Failure Mode ExtractionKeivan Rezaei, Mehrdad Saberi, Mazda Moayeri, Soheil FeiziICLR 2024 · 9 citations
- Explaining mispredictions of machine learning models using rule inductionJürgen Cito, Isil Dillig, Seohyun Kim, Vijayaraghavan Murali et al.FSE 2021 · 26 citations
- Goal Driven Discovery of Distributional Differences via Language DescriptionsRuiqi Zhong, Peter Zhang, Steve Li, Jinwoo Ahn et al.NeurIPS 2023 · 81 citations
- Mining Long Tail Bugs: Identifying Rare and Overlooked Issues in CodeWentao Liang, Yanjun Wu, Xiang Ling, Tianyue Luo et al.FSE 2026
