Predicting Rare Events by Shrinking Towards Proportional Odds
Gregory Faletto, Jacob Bien
Abstract
Training classifiers is difficult with severe class imbalance, but many rare events are the culmination of a sequence with much more common intermediate outcomes. For example, in online marketing a user first sees an ad, then may click on it, and finally may make a purchase; estimating the probability of purchases is difficult because of their rarity. We show both theoretically and through data experiments that the more abundant data in earlier steps may be leveraged to improve estimation of probabilities of rare events. We present PRESTO, a relaxation of the proportional odds model for ordinal regression. Instead of estimating weights for one separating hyperplane that is shifted by separate intercepts for each of the estimated Bayes decision boundaries between adjacent pairs of categorical responses, we estimate separate weights for each of these transitions. We impose an L1 penalty on the differences between weights for the same feature in adjacent weight vectors in order to shrink towards the proportional odds model. We prove that PRESTO consistently estimates the decision boundary weights under a sparsity assumption. Synthetic and real data experiments show that our method can estimate rare probabilities in this setting better than both logistic regression on the rare category, which fails to borrow strength from more abundant categories, and the proportional odds model, which is too inflexible.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bd910df9-8d1b-4477-aa88-c6b2570b600bCited by top-tier papers1
Ask how each one uses itRelated papers
- Ord2Seq: Regarding Ordinal Regression as Label Sequence PredictionJinhong Wang, Yi Cheng, Jintai Chen, Tingting Chen et al.ICCV 2023 · 18 citations
- Logistic Regression for Massive Data with Rare EventsHaiying WangICML 2020 · 29 citations
- Scale-invariant Optimal Sampling for Rare-events Data and Sparse ModelsJing Wang, HaiYing Wang, Hao ZhangNeurIPS 2024 · 1 citation
- Improving Model Probability Calibration by Integration of Large Data Sources with Biased LabelsRenat Sergazinov, Richard Chen, Cheng Ji, Jing Wu et al.AAAI 2025
- Fast OSCAR and OWL Regression via Safe Screening RulesRunxue Bao, Bin Gu, Heng HuangICML 2020 · 41 citations
