Background Splitting: Finding Rare Classes in a Sea of Background
Ravi Teja Mullapudi, Fait Poms, William R. Mark, Deva Ramanan, Kayvon Fatahalian
Abstract
We focus on the problem of training deep image classification models for a small number of extremely rare categories. In this common, real-world scenario, almost all images belong to the background category in the dataset. We find that state-of-the-art approaches for training on imbalanced datasets do not produce accurate deep models in this regime. Our solution is to split the large, visually diverse background into many smaller, visually similar categories during training. We implement this idea by extending an image classification model with an additional auxiliary loss that learns to mimic the predictions of a pre-existing classification model on the training set. The auxiliary loss requires no additional human labels and regularizes feature learning in the shared network trunk by forcing the model to discriminate between auxiliary categories for all training set examples, including those belonging to the monolithic background of the main rare category classification task. To evaluate our method we contribute modified versions of the iNaturalist and Places365 datasets where only a small subset of rare category labels are available during training (all other images are labeled as background). By jointly learning to recognize both the selected rare categories and auxiliary categories, our approach yields models that perform 8.3 mAP points higher than stateof-the-art imbalanced learning baselines when 98.30% of the data is background, and up to 42.3 mAP points higher than fine-tuning baselines when 99.98% of the data is background.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8cfc7723-f6ea-4d2f-ac80-8200702772ceCited by top-tier papers6
- Distilling Virtual Examples for Long-tailed RecognitionYin-Yin He, Jianxin Wu, Xiu-Shen WeiICCV 2021 · 129 citations
- Agile Modeling: From Concept to Classifier in MinutesOtilia Stretcu, Edward Vendrow, Kenji Hata, Krishnamurthy Viswanathan et al.ICCV 2023 · 19 citations
- Learning Rare Category Classifiers on a Tight Labeling BudgetRavi Teja Mullapudi, Fait Poms, William R. Mark, Deva Ramanan et al.ICCV 2021 · 17 citations
- Low-Bandwidth Self-Improving Transmission of Rare Training DataShilpa Anna George, Haithem Turki, Ziqiang Feng, Deva Ramanan et al.MobiCom 2023 · 5 citations
- Pairwise Maximum Likelihood For Multi-Class Logistic Regression Model With Multiple Rare ClassesXuetong Li, Danyang Huang, Hansheng WangICML 2025
Builds on4
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan et al.ICLR 2020 · 1,496 citations
- Semantic-Aware Knowledge Preservation for Zero-Shot Sketch-Based Image RetrievalQing Liu, Lingxi Xie, Huiyu Wang, Alan L. YuilleICCV 2019 · 126 citations
- DistInit: Learning Video Representations Without a Single Labeled VideoRohit Girdhar, Du Tran, Lorenzo Torresani, Deva RamananICCV 2019 · 59 citations
- Overcoming Classifier Imbalance for Long-Tail Object Detection With Balanced Group SoftmaxYu Li, Tao Wang, Bingyi Kang, Sheng Tang et al.CVPR 2020
Related papers
- Distributional Robustness Loss for Long-tail LearningDvir Samuel, Gal ChechikICCV 2021 · 128 citations
- RSG: A Simple but Effective Module for Learning Imbalanced DatasetsJianfeng Wang, Thomas Lukasiewicz, Xiaolin Hu, Jianfei Cai et al.CVPR 2021
- Retrieval Augmented Classification for Long-Tail Visual RecognitionAlexander Long, Wei Yin, Thalaiyasingam Ajanthan, Vu Nguyen et al.CVPR 2022 · 64 citations
- Multi-Label Learning From Single Positive LabelsElijah Cole, Oisin Mac Aodha, Titouan Lorieul, Pietro Perona et al.CVPR 2021
- HGLTR: Hierarchical Knowledge Injection for Calibrating Pre-trained Models in Long-Tail RecognitionJinpeng Zheng, Shao-Yuan Li, Gan Xu, Wenhai Wan et al.AAAI 2026
