Reconciling Object-Level and Global-Level Objectives for Long-Tail Detection
Shaoyu Zhang, Chen Chen, Silong Peng
Abstract
Large vocabulary object detectors are often faced with the long-tailed label distributions, seriously degrading their ability to detect rarely seen categories. On one hand, the rare objects are prone to be misclassified as frequent categories. On the other hand, due to the limitation on the total number of detections per image, detectors usually rank all the confidence scores globally and filter out the lower-ranking ones. This may result in missed detection during inference, especially for the rare categories that naturally come with lower scores. Existing methods mainly focus on the former problem and design various classification loss to enhance the object-level classification accuracy, but largely overlook the global-level ranking task. In this paper, we propose a novel framework that Reconciles Object-level and Global-level (ROG) objectives to address both problems. As a multi-task learning framework, ROG simultaneously trains the model with two tasks: classifying each object proposal individually and ranking all the confidence scores globally. Specifically, complementary to the object-level classification loss for model discrimination, we design a generalized average precision (GAP) loss to explicitly optimize the global-level score ranking across different objects. For each category, GAP loss generates balanced gradients to rectify the ranking errors. In experiments, we show that GAP loss is highly versatile to be plugged into various advanced methods and brings considerable benefits. Code is at https://github.com/EricZsy/ROG.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Long-tailed Object Detection Pretraining: Dynamic Rebalancing Contrastive Learning with Dual ReconstructionChen-Long Duan, Yong Li, Xiu-Shen Wei, Lin ZhaoNeurIPS 2024 · 9 citations
- Fractal Calibration for Long-tailed Object DetectionKonstantinos Panagiotis Alexandridis, Ismail Elezi, Jiankang Deng, Anh Nguyen et al.CVPR 2025
Builds on23
- CenterNet: Keypoint Triplets for Object DetectionKaiwen Duan, Song Bai, Lingxi Xie, Honggang Qi et al.ICCV 2019 · 3,348 citations
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan et al.ICLR 2020 · 1,496 citations
- Balanced Meta-Softmax for Long-Tailed Visual RecognitionJiawei Ren, Cunjun Yu, Shunan Sheng, Xiao Ma et al.NeurIPS 2020 · 861 citations
- Long-Tailed Classification by Keeping the Good and Removing the Bad Momentum Causal EffectKaihua Tang, Jianqiang Huang, Hanwang ZhangNeurIPS 2020 · 533 citations
- FASA: Feature Augmentation and Sampling Adaptation for Long-Tailed Instance SegmentationYuhang Zang, Chen Huang, Chen Change LoyICCV 2021 · 142 citations
Related papers
- DropLoss for Long-Tail Instance SegmentationTing-I Hsieh, Esther Robb, Hwann-Tzong Chen, Jia-Bin HuangAAAI 2021 · 53 citations
- Adaptive Class Suppression Loss for Long-Tail Object DetectionTong Wang, Yousong Zhu, Chaoyang Zhao, Wei Zeng et al.CVPR 2021
- Adaptive Hierarchical Representation Learning for Long-Tailed Object DetectionBanghuai LiCVPR 2022 · 16 citations
- Equalization Loss for Long-Tailed Object RecognitionJingru Tan, Changbao Wang, Buyu Li, Quanquan Li et al.CVPR 2020
- Overcoming Classifier Imbalance for Long-Tail Object Detection With Balanced Group SoftmaxYu Li, Tao Wang, Bingyi Kang, Sheng Tang et al.CVPR 2020
