H2FA R-CNN: Holistic and Hierarchical Feature Alignment for Cross-domain Weakly Supervised Object Detection
Yunqiu Xu, Yifan Sun, Zongxin Yang, Jiaxu Miao, Yi Yang
摘要
Cross-domain weakly supervised object detection (CD-WSOD) aims to adapt the detection model to a novel target domain with easily acquired image-level annotations. How to align the source and target domains is critical to the CDWSOD accuracy. Existing methods usually focus on partial detection components for domain alignment. In contrast, this paper considers that all the detection components are important and proposes a Holistic and Hier-archical Feature Alignment (H <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">2</sup> FA) R-CNN. H2FA R-CNN enforces two image-level alignments for the backbone features, as well as two instance-level alignments for the RPN and detection head. This coarse-to-fine aligning hierarchy is in pace with the detection pipeline, i.e., processing the image-level feature and the instance-level features from bottom to top. Importantly, we devise a novel hybrid supervision method for learning two instance-level align-ments. It enables the RPN and detection head to simultane-ously receive weak/full supervision from the target/source domains. Combining all these feature alignments, H2 FA R-CNN effectively mitigates the gap between the source and target domains. Experimental results show that H2 FA R-CNN significantly improves cross-domain object detection accuracy and sets new state of the art on popular benchmarks. Code and pre-trained models are available at https://github.com/XuYunqiu/H2FA_R-CNN.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Open-Vocabulary Object Detection via Language HierarchyJiaxing Huang, Jingyi Zhang, Kai Jiang, Shijian LuNeurIPS 2024 · 被引用 16 次
- DSD-DA: Distillation-based Source Debiasing for Domain Adaptive Object DetectionYongchao Feng, Shiwei Li, Yingjie Gao, Ziyue Huang 等ICML 2024 · 被引用 11 次
- MC-Bench: A Benchmark for Multi-Context Visual Grounding in the Era of MLLMsYunqiu Xu, Linchao Zhu, Yi YangICCV 2025 · 被引用 7 次
- TRKT: Weakly Supervised Dynamic Scene Graph Generation with Temporal-Enhanced Relation-Aware Knowledge TransferringZhu Xu, Ting Lei, Zhimin Li, Guan Wang 等ICCV 2025 · 被引用 3 次
- TITAN: Query-Token Based Domain Adaptive Adversarial LearningTajamul Ashraf, Janibul BashirICCV 2025 · 被引用 2 次
它引用的顶会 Paper35
- Objects365: A Large-Scale, High-Quality Dataset for Object DetectionShuai Shao, Zeming Li, Tianyuan Zhang, Chao Peng 等ICCV 2019 · 被引用 1,018 次
- Multi-Adversarial Faster-RCNN for Unrestricted Object DetectionZhenwei He, Lei ZhangICCV 2019 · 被引用 352 次
- A Robust Learning Approach to Domain Adaptive Object DetectionMehran Khodabandeh, Arash Vahdat, Mani Ranjbar, William G. MacreadyICCV 2019 · 被引用 273 次
- Self-Training and Adversarial Background Regularization for Unsupervised Domain Adaptive One-Stage Object DetectionSeunghyeon Kim, Jaehoon Choi, Taekyung Kim, Changick KimICCV 2019 · 被引用 211 次
- A Free Lunch for Unsupervised Domain Adaptive Object Detection without Source DataXianfeng Li, Weijie Chen, Di Xie, Shicai Yang 等AAAI 2021 · 被引用 181 次
相关 Paper
- DETR with Additional Global Aggregation for Cross-domain Weakly Supervised Object DetectionZongheng Tang, Yifan Sun, Si Liu, Yi YangCVPR 2023
- Informative and Consistent Correspondence Mining for Cross-Domain Weakly Supervised Object DetectionLuwei Hou, Yu Zhang, Kui Fu, Jia LiCVPR 2021
- Cross-domain Object Detection through Coarse-to-Fine Feature AdaptationYangtao Zheng, Di Huang, Songtao Liu, Yunhong WangCVPR 2020
- iFAN: Image-Instance Full Alignment Networks for Adaptive Object DetectionChenfan Zhuang, Xintong Han, Weilin Huang, Matthew R. ScottAAAI 2020 · 被引用 92 次
- Exploring Categorical Regularization for Domain Adaptive Object DetectionChang-Dong Xu, Xing-Ran Zhao, Xin Jin, Xiu-Shen WeiCVPR 2020
