Towards Robust Object Detection Invariant to Real-World Domain Shifts
Qi Fan, Mattia Segù, Yu-Wing Tai, Fisher Yu, Chi-Keung Tang, Bernt Schiele, Dengxin Dai
摘要
Safety-critical applications such as autonomous driving require robust object detection invariant to real-world domain shifts. Such shifts can be regarded as different domain styles, which can vary substantially due to environment changes, but deep models only know the training domain style. Such domain style gap impedes object detection generalization on diverse real-world domains. Existing classification domain generalization (DG) methods cannot effectively solve the robust object detection problem, because they either rely on multiple source domains with large style variance or destroy the content structures of the original images. In this paper, we analyze and investigate effective solutions to overcome domain style overfitting for robust object detection without the above shortcomings. Our method, dubbed as Normalization Perturbation (NP), perturbs the channel statistics of source domain low-level features to synthesize various latent styles, so that the trained deep model can perceive diverse potential domains and generalizes well even without observations of target domain data in training. This approach is motivated by the observation that feature channel statistics of the target domain images deviate around the source domain statistics. We further explore the style-sensitive channels for effective style synthesis. Normalization Perturbation only relies on a single source domain and is surprisingly simple and effective, contributing a practical solution by effectively adapting or generalizing classification DG methods to robust object detection. Extensive experiments demonstrate the effectiveness of our method for generalizing object detectors under real-world domain shifts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- PØDA: Prompt-driven Zero-shot Domain AdaptationMohammad Fahes, Tuan-Hung Vu, Andrei Bursuc, Patrick Pérez 等ICCV 2023 · 被引用 82 次
- Stronger, Fewer, & Superior: Harnessing Vision Foundation Models for Domain Generalized Semantic SegmentationZhixiang Wei, Lin Chen, Yi Jin, Xiaoxiao Ma 等CVPR 2024 · 被引用 61 次
- Domain-Rectifying Adapter for Cross-Domain Few-Shot SegmentationJiapeng Su, Qi Fan, Wenjie Pei, Guangming Lu 等CVPR 2024 · 被引用 22 次
- DARTH: Holistic Test-time Adaptation for Multiple Object TrackingMattia Segù, Bernt Schiele, Fisher YuICCV 2023 · 被引用 17 次
- MRFP: Learning Generalizable Semantic Segmentation from Sim-2-Real with Multi-Resolution Feature PerturbationSumanth Udupa, Prajwal Gurunath, Aniruddh Sikdar, Suresh SundaramCVPR 2024 · 被引用 10 次
它引用的顶会 Paper31
- An Empirical Study of Training Self-Supervised Vision TransformersXinlei Chen, Saining Xie, Kaiming HeICCV 2021 · 被引用 2,340 次
- Tent: Fully Test-Time Adaptation by Entropy MinimizationDequan Wang, Evan Shelhamer, Shaoteng Liu, Bruno A. Olshausen 等ICLR 2021 · 被引用 1,731 次
- Domain Generalization with MixStyleKaiyang Zhou, Yongxin Yang, Yu Qiao, Tao XiangICLR 2021 · 被引用 986 次
- Improving robustness against common corruptions by covariate shift adaptationSteffen Schneider, Evgenia Rusak, Luisa Eck, Oliver Bringmann 等NeurIPS 2020 · 被引用 688 次
- MEMO: Test Time Robustness via Adaptation and AugmentationMarvin Zhang, Sergey Levine, Chelsea FinnNeurIPS 2022 · 被引用 595 次
相关 Paper
- Uncertainty Modeling for Out-of-Distribution GeneralizationXiaotong Li, Yongxing Dai, Yixiao Ge, Jun Liu 等ICLR 2022 · 被引用 237 次
- DomainDrop: Suppressing Domain-Sensitive Channels for Domain GeneralizationJintao Guo, Lei Qi, Yinghuan ShiICCV 2023 · 被引用 47 次
- Rethinking Open-World Object Detection in Autonomous Driving ScenariosZeyu Ma, Yang Yang, Guoqing Wang, Xing Xu 等ACM MM 2022 · 被引用 39 次
- A Robust Learning Approach to Domain Adaptive Object DetectionMehran Khodabandeh, Arash Vahdat, Mani Ranjbar, William G. MacreadyICCV 2019 · 被引用 273 次
- Geometric and Textural Augmentation for Domain Gap ReductionXiao-Chang Liu, Yongliang Yang, Peter HallCVPR 2022 · 被引用 16 次
