PIT: Position-Invariant Transform for Cross-FoV Domain Adaptation
Qiqi Gu, Qianyu Zhou, Minghao Xu, Zhengyang Feng, Guangliang Cheng, Xuequan Lu, Jianping Shi, Lizhuang Ma
Abstract
Cross-domain object detection and semantic segmentation have witnessed impressive progress recently. Existing approaches mainly consider the domain shift resulting from external environments including the changes of background, illumination or weather, while distinct camera intrinsic parameters appear commonly in different domains and their influence for domain adaptation has been very rarely explored. In this paper, we observe that the Field of View (FoV) gap induces noticeable instance appearance differences between the source and target domains. We further discover that the FoV gap between two domains impairs domain adaptation performance under both the FoV-increasing (source FoV < target FoV) and FoV-decreasing cases. Motivated by the observations, we propose the Position-Invariant Transform (PIT) to better align images in different domains. We also introduce a reverse PIT for mapping the transformed/aligned images back to the original image space, and design a loss re-weighting strategy to accelerate the training process. Our method can be easily plugged into existing crossdomain detection/segmentation frameworks, while bringing about negligible computational overhead. Extensive experiments demonstrate that our method can soundly boost the performance on both cross-domain object detection and segmentation for state-of-the-art techniques. Our code is available at https://github.com/sheepooo/ PIT-Position-Invariant-Transform .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5e06eee5-136b-4092-bf1b-c080e98db0bcCited by top-tier papers18
- Bending Reality: Distortion-aware Transformers for Adapting to Panoramic Semantic SegmentationJiaming Zhang, Kailun Yang, Chaoxiang Ma, Simon Reiß et al.CVPR 2022 · 100 citations
- Look at the Neighbor: Distortion-aware Unsupervised Domain Adaptation for Panoramic Semantic SegmentationXu Zheng, Tianbo Pan, Yunhao Luo, Lin WangICCV 2023 · 46 citations
- H2FA R-CNN: Holistic and Hierarchical Feature Alignment for Cross-domain Weakly Supervised Object DetectionYunqiu Xu, Yifan Sun, Zongxin Yang, Jiaxu Miao et al.CVPR 2022 · 40 citations
- Revisiting Domain-Adaptive 3D Object Detection by Reliable, Diverse and Class-balanced Pseudo-LabelingZhuoxiao Chen, Yadan Luo, Zheng Wang, Mahsa Baktashmotlagh et al.ICCV 2023 · 40 citations
- BA-SAM: Scalable Bias-Mode Attention Mask for Segment Anything ModelYiran Song, Qianyu Zhou, Xiangtai Li, Deng-Ping Fan et al.CVPR 2024 · 19 citations
Builds on16
- Confidence Regularized Self-TrainingYang Zou, Zhiding Yu, Xiaofeng Liu, B. V. K. Vijaya Kumar et al.ICCV 2019 · 901 citations
- Semi-Supervised Domain Adaptation via Minimax EntropyKuniaki Saito, Donghyun Kim, Stan Sclaroff, Trevor Darrell et al.ICCV 2019 · 725 citations
- Adversarial Domain Adaptation with Domain MixupMinghao Xu, Jian Zhang, Bingbing Ni, Teng Li et al.AAAI 2020 · 499 citations
- Multi-Adversarial Faster-RCNN for Unrestricted Object DetectionZhenwei He, Lei ZhangICCV 2019 · 352 citations
- Domain Adaptation for Structured Output via Discriminative Patch RepresentationsYi-Hsuan Tsai, Kihyuk Sohn, Samuel Schulter, Manmohan ChandrakerICCV 2019 · 333 citations
Related papers
- Towards Domain Generalization for Multi-view 3D Object Detection in Bird-Eye-ViewShuo Wang, Xinhai Zhao, Hai-Ming Xu, Zehui Chen et al.CVPR 2023
- Knowledge Mining and Transferring for Domain Adaptive Object DetectionKun Tian, Chenghao Zhang, Ying Wang, Shiming Xiang et al.ICCV 2021 · 54 citations
- Exploring Categorical Regularization for Domain Adaptive Object DetectionChang-Dong Xu, Xing-Ran Zhao, Xin Jin, Xiu-Shen WeiCVPR 2020
- Cross-Domain Semantic Segmentation via Domain-Invariant Interactive Relation TransferFengmao Lv, Tao Liang, Xiang Chen, Guosheng LinCVPR 2020
- Rethinking Open-World Object Detection in Autonomous Driving ScenariosZeyu Ma, Yang Yang, Guoqing Wang, Xing Xu et al.ACM MM 2022 · 39 citations
