SIOD: Single Instance Annotated Per Category Per Image for Object Detection
Hanjun Li, Xingjia Pan, Ke Yan, Fan Tang, Wei-Shi Zheng
Abstract
Object detection under imperfect data receives great attention recently. Weakly supervised object detection (WSOD) suffers from severe localization issues due to the lack of instance-level annotation, while semi-supervised object detection (SSOD) remains challenging led by the inter-image discrepancy between labeled and unlabeled data. In this study, we propose the Single Instance annotated Object Detection (SIOD), requiring only one instance annotation for each existing category in an image. Degraded from inter-task (WSOD) or inter-image (SSOD) discrepancies to the intra-image discrepancy, SIOD provides more reliable and rich prior knowledge for mining the rest of unlabeled instances and trades off the annotation cost and performance. Under the SIOD setting, we propose a simple yet effective framework, termed Dual-Mining (DMiner), which consists of a Similarity-based Pseudo Label Generating module (SPLG) and a Pixel-level Group Contrastive Learning module (PGCL). SPLG firstly mines latent instances from feature representation space to alleviate the annotation missing problem. To avoid being misled by inaccurate pseudo labels, we propose PGCL to boost the tolerance to false pseudo labels. Extensive experiments on MS COCO verify the feasibility of the SIOD setting and the superiority of the proposed method, which obtains consistent and significant improvements compared to baseline methods and achieves comparable results with fully supervised object detection (FSOD) methods with only 40% instances annotated. Code is available at https: //github.com/solicucu/SIOD .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f33c44ce-504e-4bef-ad94-50c8c4994fd5Cited by top-tier papers4
- CoIn: Contrastive Instance Feature Mining for Outdoor 3D Object Detection with Very Limited AnnotationsQiming Xia, Jinhao Deng, Chenglu Wen, Hai Wu et al.ICCV 2023 · 34 citations
- D3G: Exploring Gaussian Prior for Temporal Sentence Grounding with Glance AnnotationHanjun Li, Xiujun Shu, Sunan He, Ruizhi Qiao et al.ICCV 2023 · 21 citations
- SparseDet: Improving Sparsely Annotated Object Detection with Pseudo-positive MiningSaksham Suri, Sai Saketh Rambhatla, Rama Chellappa, Abhinav ShrivastavaICCV 2023 · 19 citations
- Learning Class Prototypes for Unified Sparse-Supervised 3D Object DetectionYun Zhu, Le Hui, Hang Yang, Jianjun Qian et al.CVPR 2025
Builds on20
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 6,042 citations
- RepPoints: Point Set Representation for Object DetectionZe Yang, Shaohui Liu, Han Hu, Liwei Wang et al.ICCV 2019 · 1,056 citations
Related papers
- Group R-CNN for Weakly Semi-supervised Object Detection with PointsShilong Zhang, Zhuoran Yu, Liyang Liu, Xinjiang Wang et al.CVPR 2022 · 51 citations
- Mixed Supervision for Instance Learning in Object Detection with Few-shot AnnotationYi Zhong, Chengyao Wang, Shiyong Li, Zhu Zhou et al.ACM MM 2022 · 1 citation
- Instant-Teaching: An End-to-End Semi-Supervised Object Detection FrameworkQiang Zhou, Chaohui Yu, Zhibin Wang, Qi Qian et al.CVPR 2021
- Co-mining: Self-Supervised Learning for Sparsely Annotated Object DetectionTiancai Wang, Tong Yang, Jiale Cao, Xiangyu ZhangAAAI 2021 · 57 citations
- SS3D: Sparsely-Supervised 3D Object Detection from Point CloudChuandong Liu, Chenqiang Gao, Fangcen Liu, Jiang Liu et al.CVPR 2022 · 32 citations
