GA-DAN: Geometry-Aware Domain Adaptation Network for Scene Text Detection and Recognition
Fangneng Zhan, Chuhui Xue, Shijian Lu
Abstract
Recent adversarial learning research has achieved very impressive progress for modelling cross-domain data shifts in appearance space but its counterpart in modelling cross-domain shifts in geometry space lags far behind. This paper presents an innovative Geometry-Aware Domain Adaptation Network (GA-DAN) that is capable of modelling cross-domain shifts concurrently in both geometry space and appearance space and realistically converting images across domains with very different characteristics. In the proposed GA-DAN, a novel multi-modal spatial learning structure is designed which can convert a source-domain image into multiple images of different spatial views as in the target domain. A new disentangled cycle-consistency loss is introduced which balances the cycle consistency and greatly improves the concurrent learning in both appearance and geometry spaces. The proposed GA-DAN has been evaluated for the classic scene text detection and recognition tasks, and experiments show that the domain-adapted images achieve superior scene text detection and recognition performance while applied to network training.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers13
- Model Adaptation: Historical Contrastive Learning for Unsupervised Domain Adaptation without Source DataJiaxing Huang, Dayan Guan, Aoran Xiao, Shijian LuNeurIPS 2021 · 301 citations
- Diverse Image Inpainting with Bidirectional and Autoregressive TransformersYingchen Yu, Fangneng Zhan, Rongliang Wu, Jianxiong Pan et al.ACM MM 2021 · 153 citations
- Unsupervised Domain Adaptive 3D Detection with Multi-Level ConsistencyZhipeng Luo, Zhongang Cai, Changqing Zhou, Gongjie Zhang et al.ICCV 2021 · 92 citations
- RDA: Robust Domain Adaptation via Fourier Adversarial AttackingJiaxing Huang, Dayan Guan, Aoran Xiao, Shijian LuICCV 2021 · 85 citations
- EMLight: Lighting Estimation via Spherical Distribution ApproximationFangneng Zhan, Changgong Zhang, Yingchen Yu, Yuan Chang et al.AAAI 2021 · 73 citations
Related papers
- Multimodal Structure-Consistent Image-to-Image TranslationChe-Tsung Lin, Yen-Yi Wu, Po-Hao Hsu, Shang-Hong LaiAAAI 2020 · 24 citations
- Geometry-Aware Network for Domain Adaptive Semantic SegmentationYinghong Liao, Wending Zhou, Xu Yan, Zhen Li et al.AAAI 2023 · 9 citations
- Multi-Source Domain Adaptation for Visual Sentiment ClassificationChuang Lin, Sicheng Zhao, Lei Meng, Tat-Seng ChuaAAAI 2020 · 78 citations
- Geometry Normalization Networks for Accurate Scene Text DetectionJiaqi Duan, Youjiang Xu, Zhanghui Kuang, Xiaoyu Yue et al.ICCV 2019 · 33 citations
- DRANet: Disentangling Representation and Adaptation Networks for Unsupervised Cross-Domain AdaptationSeunghun Lee, Sunghyun Cho, Sunghoon ImCVPR 2021
