UniT: Unified Knowledge Transfer for Any-Shot Object Detection and Segmentation
Siddhesh Khandelwal, Raghav Goyal, Leonid Sigal
Abstract
Methods for object detection and segmentation rely on large scale instance-level annotations for training, which are difficult and time-consuming to collect. Efforts to alleviate this look at varying degrees and quality of supervision. Weakly-supervised approaches draw on image-level labels to build detectors/segmentors, while zero/few-shot methods assume abundant instance-level data for a set of base classes, and none to a few examples for novel classes. This taxonomy has largely siloed algorithmic designs. In this work, we aim to bridge this divide by proposing an intuitive and unified semi-supervised model that is applicable to a range of supervision: from zero to a few instance-level samples per novel class. For base classes, our model learns a mapping from weakly-supervised to fully-supervised detectors/segmentors. By learning and leveraging visual and lingual similarities between the novel and base classes, we transfer those mappings to obtain detectors/segmentors for novel classes; refining them with a few novel class instance-level annotated samples, if available. The overall model is end-to-end trainable and highly flexible. Through extensive experiments on MS-COCO [32] and Pascal VOC [14] benchmark datasets we show improved performance in a variety of settings.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bc97282f-9d90-46d7-9b17-33e0bcc83bcbCited by top-tier papers5
- Feature-Proxy Transformer for Few-Shot SegmentationJian-Wei Zhang, Yifan Sun, Yi Yang, Wei ChenNeurIPS 2022 · 105 citations
- Decoupled Adaptation for Cross-Domain Object DetectionJunguang Jiang, Baixu Chen, Jianmin Wang, Mingsheng LongICLR 2022 · 88 citations
- H2FA R-CNN: Holistic and Hierarchical Feature Alignment for Cross-domain Weakly Supervised Object DetectionYunqiu Xu, Yifan Sun, Zongxin Yang, Jiaxu Miao et al.CVPR 2022 · 40 citations
- Segmentation-grounded Scene Graph GenerationSiddhesh Khandelwal, Mohammed Suhail, Leonid SigalICCV 2021 · 32 citations
- Open-Vocabulary Object Detection via Language HierarchyJiaxing Huang, Jingyi Zhang, Kai Jiang, Shijian LuNeurIPS 2024 · 16 citations
Builds on10
- Few-Shot Object Detection via Feature ReweightingBingyi Kang, Zhuang Liu, Xin Wang, Fisher Yu et al.ICCV 2019 · 835 citations
- Frustratingly Simple Few-Shot Object DetectionXin Wang, Thomas E. Huang, Joseph Gonzalez, Trevor Darrell et al.ICML 2020 · 723 citations
- Meta R-CNN: Towards General Solver for Instance-Level Low-Shot LearningXiaopeng Yan, Ziliang Chen, Anni Xu, Xiaoxi Wang et al.ICCV 2019 · 590 citations
- Meta-Learning to Detect Rare ObjectsYu-Xiong Wang, Deva Ramanan, Martial HebertICCV 2019 · 339 citations
- Zero-Shot Grounding of Objects From Natural Language QueriesArka Sadhu, Kan Chen, Ram NevatiaICCV 2019 · 176 citations
Related papers
- CaT: Weakly Supervised Object Detection with Category TransferTianyue Cao, Lianyu Du, Xiaoyun Zhang, Siheng Chen et al.ICCV 2021 · 22 citations
- Open-Vocabulary Object Detection Using CaptionsAlireza Zareian, Kevin Dela Rosa, Derek Hao Hu, Shih-Fu ChangCVPR 2021
- Mixed Supervised Object Detection by Transferring Mask Prior and Semantic SimilarityYan Liu, Zhijie Zhang, Li Niu, Junjie Chen et al.NeurIPS 2021 · 25 citations
- Parallel Detection-and-Segmentation Learning for Weakly Supervised Instance SegmentationYunhang Shen, Liujuan Cao, Zhiwei Chen, Baochang Zhang et al.ICCV 2021 · 22 citations
- ContrastMask: Contrastive Learning to Segment Every ThingXuehui Wang, Kai Zhao, Ruixin Zhang, Shouhong Ding et al.CVPR 2022 · 45 citations
