Generative Modeling for Small-Data Object Detection
Lanlan Liu, Michael Muelly, Jia Deng, Tomas Pfister, Li-Jia Li
摘要
This paper explores object detection in the small data regime, where only a limited number of annotated bounding boxes are available due to data rarity and annotation expense. This is a common challenge today with machine learning being applied to many new tasks where obtaining training data is more challenging, e.g. in medical images with rare diseases that doctors sometimes only see once in their life-time. In this work we explore this problem from a generative modeling perspective by learning to generate new images with associated bounding boxes, and using these for training an object detector. We show that simply training previously proposed generative models does not yield satisfactory performance due to them optimizing for image realism rather than object detection accuracy. To this end we develop a new model with a novel unrolling mechanism that jointly optimizes the generative model and a detector such that the generated images improve the performance of the detector. We show this method outperforms the state of the art on two challenging datasets, disease detection and small data pedestrian detection, improving the average precision on NIH Chest X-ray by a relative 20% and localization accuracy by a relative 50%. * This work was conducted when Lanlan Liu was an intern at Google.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- FASA: Feature Augmentation and Sampling Adaptation for Long-Tailed Instance SegmentationYuhang Zang, Chen Huang, Chen Change LoyICCV 2021 · 被引用 142 次
- Artificial Dummies for Urban Dataset AugmentationAntonín Vobecký, David Hurych, Michal Uricár, Patrick Pérez 等AAAI 2021 · 被引用 17 次
- StarNet: towards Weakly Supervised Few-Shot Object DetectionLeonid Karlinsky, Joseph Shtok, Amit Alfassy, Moshe Lichtenstein 等AAAI 2021 · 被引用 17 次
- Box-based Refinement for Weakly Supervised and Unsupervised Localization TasksEyal Gomel, Tal Shaharabany, Lior WolfICCV 2023 · 被引用 6 次
相关 Paper
- Unsupervised Object Detection Pretraining with Joint Object Priors Generation and Detector LearningYizhou Wang, Meilin Chen, Shixiang Tang, Feng Zhu 等NeurIPS 2022 · 被引用 2 次
- Pix2seq: A Language Modeling Framework for Object DetectionTing Chen, Saurabh Saxena, Lala Li, David J. Fleet 等ICLR 2022 · 被引用 435 次
- Open-Det: An Efficient Learning Framework for Open-Ended DetectionGuiping Cao, Tao Wang, Wenjian Huang, Xiangyuan Lan 等ICML 2025
- MedGR2: Breaking the Data Barrier for Medical Reasoning via Generative Reward LearningWeihai Zhi, Jiayan Guo, Shangyang LiAAAI 2026 · 被引用 5 次
- Tell Me What They're Holding: Weakly-Supervised Object Detection with Transferable Knowledge from Human-Object InteractionDaesik Kim, Gyujeong Lee, Jisoo Jeong, Nojun KwakAAAI 2020 · 被引用 16 次
