Long-tailed Object Detection Pretraining: Dynamic Rebalancing Contrastive Learning with Dual Reconstruction
Chen-Long Duan, Yong Li, Xiu-Shen Wei, Lin Zhao
摘要
Pre-training plays a vital role in various vision tasks, such as object recognition and detection. Commonly used pre-training methods, which typically rely on randomized approaches like uniform or Gaussian distributions to initialize model parameters, often fall short when confronted with long-tailed distributions, especially in detection tasks. This is largely due to extreme data imbalance and the issue of simplicity bias. In this paper, we introduce a novel pre-training framework for object detection, called Dynamic Rebalancing Contrastive Learning with Dual Reconstruction (2DRCL). Our method builds on a Holistic-Local Contrastive Learning mechanism, which aligns pre-training with object detection by capturing both global contextual semantics and detailed local patterns. To tackle the imbalance inherent in long-tailed data, we design a dynamic rebalancing strategy that adjusts the sampling of underrepresented instances throughout the pre-training process, ensuring better representation of tail classes. Moreover, Dual Reconstruction addresses simplicity bias by enforcing a reconstruction task aligned with the self-consistency principle, specifically benefiting underrepresented tail classes. Experiments on COCO and LVIS v1.0 datasets demonstrate the effectiveness of our method, particularly in improving the mAP/AP scores for tail classes.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper30
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi 等NeurIPS 2020 · 被引用 2,611 次
- The Pitfalls of Simplicity Bias in Neural NetworksHarshay Shah, Kaustav Tamuly, Aditi Raghunathan, Prateek Jain 等NeurIPS 2020 · 被引用 503 次
- Self-Supervised Aggregation of Diverse Experts for Test-Agnostic Long-Tailed RecognitionYifan Zhang, Bryan Hooi, Lanqing Hong, Jiashi FengNeurIPS 2022 · 被引用 214 次
相关 Paper
- Boosting Long-tailed Object Detection via Step-wise Learning on Smooth-tail DataNa Dong, Yongqiang Zhang, Mingli Ding, Gim Hee LeeICCV 2023 · 被引用 8 次
- On the Effectiveness of Out-of-Distribution Data in Self-Supervised Long-Tail LearningJianhong Bai, Zuozhu Liu, Hualiang Wang, Jin Hao 等ICLR 2023 · 被引用 6 次
- Adaptive Hierarchical Representation Learning for Long-Tailed Object DetectionBanghuai LiCVPR 2022 · 被引用 16 次
- Equalization Loss v2: A New Gradient Balance Approach for Long-Tailed Object DetectionJingru Tan, Xin Lu, Gang Zhang, Changqing Yin 等CVPR 2021
- SimLTD: Simple Supervised and Semi-Supervised Long-Tailed Object DetectionPhi Vu TranCVPR 2025
