Enhancing Simple Models by Exploiting What They Already Know
Amit Dhurandhar, Karthikeyan Shanmugam, Ronny Luss
Abstract
There has been recent interest in improving performance of simple models for multiple reasons such as interpretability, robust learning from small data, deployment in memory constrained settings as well as environmental considerations. In this paper, we propose a novel method SRatio that can utilize information from high performing complex models (viz. deep neural networks, boosted trees, random forests) to reweight a training dataset for a potentially low performing simple model of much lower complexity such as a decision tree or a shallow network enhancing its performance. Our method also leverages the per sample hardness estimate of the simple model which is not the case with the prior works which primarily consider the complex model's confidences/predictions and is thus conceptually novel. Moreover, we generalize and formalize the concept of attaching probes to intermediate layers of a neural network to other commonly used classifiers and incorporate this into our method. The benefit of these contributions is witnessed in the experiments where on 6 UCI datasets and CIFAR-10 we outperform competitors in a majority (16 out of 27) of the cases and tie for best performance in the remaining cases. In fact, in a couple of cases, we even approach the complex model's performance. We also conduct further experiments to validate assertions and intuitively understand why our method works. Theoretically, we motivate our approach by showing that the weighted loss minimized by simple models using our weighting upper bounds the loss of the complex model.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fcf45b46-9172-47e5-adb1-57a0fd834a68Cited by top-tier papers2
- Locally Invariant Explanations: Towards Stable and Unidirectional Explanations through Local Invariant LearningAmit Dhurandhar, Karthikeyan Natesan Ramamurthy, Kartik Ahuja, Vijay AryaNeurIPS 2023 · 7 citations
- Auto-Transfer: Learning to Route Transferable RepresentationsKeerthiram Murugesan, Vijay Sadashivaiah, Ronny Luss, Karthikeyan Shanmugam et al.ICLR 2022 · 6 citations
Related papers
- Symbolic Regression Enhanced Decision Trees for Classification TasksKei Sen Fong, Mehul MotaniAAAI 2024 · 12 citations
- Revive Re-weighting in Imbalanced Learning by Density Ratio EstimationJiaan Luo, Feng Hong, Jiangchao Yao, Bo Han et al.NeurIPS 2024 · 16 citations
- Robust Model Compression Using Deep HypothesesOmri Armstrong, Ran Gilad-BachrachAAAI 2021 · 2 citations
- Distilling Cross-Task Knowledge via Relationship MatchingHan-Jia Ye, Su Lu, De-Chuan ZhanCVPR 2020
- Estimating informativeness of samples with Smooth Unique InformationHrayr Harutyunyan, Alessandro Achille, Giovanni Paolini, Orchid Majumder et al.ICLR 2021 · 26 citations
