Sylph: A Hypernetwork Framework for Incremental Few-shot Object Detection
Li Yin, Juan M. Perez-Rua, Kevin J. Liang
Abstract
We study the challenging incremental few-shot object de-tection (iFSD) setting. Recently, hypernetwork-based approaches have been studied in the context of continuous and finetune-free iFSD with limited success. We take a closer look at important design choices of such methods, leading to several key improvements and resulting in a more accurate and flexible framework, which we call Sylph. In particular, we demonstrate the effectiveness of decou-pling object classification from localization by leveraging a base detector that is pretrained for class-agnostic local-ization on large-scale dataset. Contrary to what previous results have suggested, we show that with a carefully de-signed class-conditional hypernetwork, finetune-free iFSD can be highly effective, especially when a large number of base categories with abundant data are available for meta-training, almost approaching alternatives that undergo test-time-training. This result is even more significant considering its many practical advantages: (1) incrementally learning new classes in sequence without additional training, (2) detecting both novel and seen classes in a single pass, and (3) no forgetting of previously seen classes. We benchmark our model on both COCO and LVIS, reporting as high as 17% AP on the long-tail rare classes on LVIS, indicating the promise of hypernetwork-based iFSD.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext de06e12c-b27a-4900-b451-975cad904456Cited by top-tier papers9
- EgoLoc: Revisiting 3D Object Localization from Egocentric Videos with Visual QueriesJinjie Mai, Abdullah Hamdi, Silvio Giancola, Chen Zhao et al.ICCV 2023 · 26 citations
- Zero-shot Generalizable Incremental Learning for Vision-Language Object DetectionJieren Deng, Haojian Zhang, Kun Ding, Jianhua Hu et al.NeurIPS 2024 · 22 citations
- PS-TTL: Prototype-based Soft-labels and Test-Time Learning for Few-shot Object DetectionYingjie Gao, Yanan Zhang, Ziyue Huang, Nanqing Liu et al.ACM MM 2024 · 13 citations
- HyperNetwork-based Decoupling to Improve Model Generalization for Few-Shot Relation ExtractionLiang Zhang, Chulun Zhou, Fandong Meng, Jinsong Su et al.EMNLP 2023 · 3 citations
- Few-Shot Incremental 3D Object Detection in Dynamic Indoor EnvironmentsYun Zhu, Jianjun Qian, Jian Yang, Jin Xie et al.CVPR 2026 · 2 citations
Builds on17
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 6,042 citations
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- CenterNet: Keypoint Triplets for Object DetectionKaiwen Duan, Song Bai, Lingxi Xie, Honggang Qi et al.ICCV 2019 · 3,348 citations
- Test-Time Training with Self-Supervision for Generalization under Distribution ShiftsYu Sun, Xiaolong Wang, Zhuang Liu, John Miller et al.ICML 2020 · 1,220 citations
- Few-Shot Object Detection via Feature ReweightingBingyi Kang, Zhuang Liu, Xin Wang, Fisher Yu et al.ICCV 2019 · 835 citations
Related papers
- Meta-Learning to Detect Rare ObjectsYu-Xiong Wang, Deva Ramanan, Martial HebertICCV 2019 · 339 citations
- Incremental Few Shot Semantic Segmentation via Class-agnostic Mask Proposal and Language-driven ClassifierLeo Shan, Wenzhang Zhou, Grace ZhaoACM MM 2023 · 20 citations
- Incremental Few-Shot Object DetectionJuan-Manuel Pérez-Rúa, Xiatian Zhu, Timothy M. Hospedales, Tao XiangCVPR 2020
- Incremental-DETR: Incremental Few-Shot Object Detection via Self-Supervised LearningNa Dong, Yongqiang Zhang, Mingli Ding, Gim Hee LeeAAAI 2023 · 54 citations
- Frustratingly Simple Few-Shot Object DetectionXin Wang, Thomas E. Huang, Joseph Gonzalez, Trevor Darrell et al.ICML 2020 · 723 citations
