Zero-Shot Aerial Object Detection with Visual Description Regularization
Zhengqing Zang, Chenyu Lin, Chenwei Tang, Tao Wang, Jiancheng Lv
摘要
Existing object detection models are mainly trained on large-scale labeled datasets. However, annotating data for novel aerial object classes is expensive since it is time-consuming and may require expert knowledge. Thus, it is desirable to study label-efficient object detection methods on aerial images. In this work, we propose a zero-shot method for aerial object detection named visual Description Regularization, or DescReg. Concretely, we identify the weak semantic-visual correlation of the aerial objects and aim to address the challenge with prior descriptions of their visual appearance. Instead of directly encoding the descriptions into class embedding space which suffers from the representation gap problem, we propose to infuse the prior inter-class visual similarity conveyed in the descriptions into the embedding learning. The infusion process is accomplished with a newly designed similarity-aware triplet loss which incorporates structured regularization on the representation space. We conduct extensive experiments with three challenging aerial object detection datasets, including DIOR, xView, and DOTA. The results demonstrate that DescReg significantly outperforms the state-of-the-art ZSD methods with complex projection designs and generative frameworks, e.g., DescReg outperforms best reported ZSD method on DIOR by 4.5 mAP on unseen classes and 8.1 in HM. We further show the generalizability of DescReg by integrating it into generative ZSD methods as well as varying the detection architecture. Codes will be released at https://github.com/zq-zang/DescReg.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- InstructSAM: A Training-free Framework for Instruction-Oriented Remote Sensing Object RecognitionYijie Zheng, Weijie Wu, Qingyun Li, Xuehui Wang 等NeurIPS 2025 · 被引用 12 次
- Self-Prompting Analogical Reasoning for UAV Object DetectionNianxin Li, Mao Ye, Lihua Zhou, Song Tang 等AAAI 2025 · 被引用 10 次
- OpenRSD: Towards Open-Prompts for Object Detection in Remote Sensing ImagesZiyue Huang, Yongchao Feng, Ziqi Liu, Shuai Yang 等ICCV 2025 · 被引用 3 次
- Open-Text Aerial Detection: A Unified Framework For Aerial Visual Grounding And DetectionGuoting Wei, Xia Yuan, Yangzhou, Haizhao Jing 等ICML 2026 · 被引用 2 次
- Dual Prompt Learning for Adapting Vision-Language Models to Downstream Image-Text RetrievalYifan Wang, Tao Wang, Chenwei Tang, Caiyang Yu 等ACM MM 2025 · 被引用 1 次
它引用的顶会 Paper11
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Few-Shot Object Detection via Feature ReweightingBingyi Kang, Zhuang Liu, Xin Wang, Fisher Yu 等ICCV 2019 · 被引用 835 次
- Oriented RepPoints for Aerial Object DetectionWentong Li, Yijie Chen, Kaixuan Hu, Jianke ZhuCVPR 2022 · 被引用 487 次
- QueryDet: Cascaded Sparse Query for Accelerating High-Resolution Small Object DetectionChenhongyi Yang, Zehao Huang, Naiyan WangCVPR 2022 · 被引用 472 次
- Clustered Object Detection in Aerial ImagesFan Yang, Heng Fan, Peng Chu, Erik Blasch 等ICCV 2019 · 被引用 384 次
相关 Paper
- Zero-Shot Object Detection by Semantics-Aware DETR with Adaptive Contrastive LossHuan Liu, Lu Zhang, Jihong Guan, Shuigeng ZhouACM MM 2023 · 被引用 6 次
- ZoRI: Towards Discriminative Zero-Shot Remote Sensing Instance SegmentationShiqi Huang, Shuting He, Bihan WenAAAI 2025
- Robust Region Feature Synthesizer for Zero-Shot Object DetectionPeiliang Huang, Junwei Han, De Cheng, Dingwen ZhangCVPR 2022 · 被引用 50 次
- SOOD: Towards Semi-Supervised Oriented Object DetectionWei Hua, Dingkang Liang, Jingyu Li, Xiaolong Liu 等CVPR 2023
- ReDet: A Rotation-Equivariant Detector for Aerial Object DetectionJiaming Han, Jian Ding, Nan Xue, Gui-Song XiaCVPR 2021
