Generating Features with Increased Crop-Related Diversity for Few-Shot Object Detection
Jingyi Xu, Hieu Le, Dimitris Samaras
摘要
Two-stage object detectors generate object proposals and classify them to detect objects in images. These proposals often do not contain the objects perfectly but overlap with them in many possible ways, exhibiting great variability in the difficulty levels of the proposals. Training a robust classifier against this crop-related variability requires abundant training data, which is not available in few-shot settings. To mitigate this issue, we propose a novel variational autoencoder (VAE) based data generation model, which is capable of generating data with increased croprelated diversity. The main idea is to transform the latent space such latent codes with different norms represent different crop-related variations. This allows us to generate features with increased crop-related diversity in difficulty levels by simply varying the latent norm. In particular, each latent code is rescaled such that its norm linearly correlates with the IoU score of the input crop w.r.t. the ground-truth box. Here the IoU score is a proxy that represents the difficulty level of the crop. We train this VAE model on base classes conditioned on the semantic code of each class and then use the trained model to generate features for novel classes. In our experiments our generated features consistently improve state-of-the-art few-shot object detection methods on the PASCAL VOC and MS COCO datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- SNIDA: Unlocking Few-Shot Object Detection with Non-Linear Semantic Decoupling AugmentationYanjie Wang, Xu Zou, Luxin Yan, Sheng Zhong 等CVPR 2024 · 被引用 22 次
- Visual Textualization for Image Prompted Object DetectionYongjian Wu, Yang Zhou, Jiya Saiyin, Bingzheng Wei 等ICCV 2025 · 被引用 1 次
- Few-shot Personalized Scanpath PredictionRuoyu Xue, Jingyi Xu, Sounak Mondal, Hieu Le 等CVPR 2025
- As Pseudo-Label Free as Possible: Leveraging Adaptive Feature Generation for Sparsely Annotated Object DetectionShuilian Yao, Yu Liu, Qi Jia, Sihong Chen 等AAAI 2025
- Few-Shot Object Detection with Foundation ModelsGuangxing Han, Ser-Nam LimCVPR 2024
它引用的顶会 Paper26
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Few-Shot Object Detection via Feature ReweightingBingyi Kang, Zhuang Liu, Xin Wang, Fisher Yu 等ICCV 2019 · 被引用 835 次
- Frustratingly Simple Few-Shot Object DetectionXin Wang, Thomas E. Huang, Joseph Gonzalez, Trevor Darrell 等ICML 2020 · 被引用 723 次
- Meta R-CNN: Towards General Solver for Instance-Level Low-Shot LearningXiaopeng Yan, Ziliang Chen, Anni Xu, Xiaoxi Wang 等ICCV 2019 · 被引用 590 次
- Meta-Learning to Detect Rare ObjectsYu-Xiong Wang, Deva Ramanan, Martial HebertICCV 2019 · 被引用 339 次
相关 Paper
- Few-Shot Object Detection via Variational Feature AggregationJiaming Han, Yuqiang Ren, Jian Ding, Ke Yan 等AAAI 2023 · 被引用 135 次
- SCHA-VAE: Hierarchical Context Aggregation for Few-Shot GenerationGiorgio Giannone, Ole WintherICML 2022 · 被引用 11 次
- Generating Representative Samples for Few-Shot ClassificationJingyi Xu, Hieu LeCVPR 2022 · 被引用 96 次
- Prototypical Variational Autoencoder for 3D Few-shot Object DetectionWeiliang Tang, Biqi Yang, Xianzhi Li, Yun-Hui Liu 等NeurIPS 2023 · 被引用 8 次
- FSCE: Few-Shot Object Detection via Contrastive Proposal EncodingBo Sun, Banghuai Li, Shengcai Cai, Ye Yuan 等CVPR 2021
