GA-DAN: Geometry-Aware Domain Adaptation Network for Scene Text Detection and Recognition
Fangneng Zhan, Chuhui Xue, Shijian Lu
摘要
Recent adversarial learning research has achieved very impressive progress for modelling cross-domain data shifts in appearance space but its counterpart in modelling cross-domain shifts in geometry space lags far behind. This paper presents an innovative Geometry-Aware Domain Adaptation Network (GA-DAN) that is capable of modelling cross-domain shifts concurrently in both geometry space and appearance space and realistically converting images across domains with very different characteristics. In the proposed GA-DAN, a novel multi-modal spatial learning structure is designed which can convert a source-domain image into multiple images of different spatial views as in the target domain. A new disentangled cycle-consistency loss is introduced which balances the cycle consistency and greatly improves the concurrent learning in both appearance and geometry spaces. The proposed GA-DAN has been evaluated for the classic scene text detection and recognition tasks, and experiments show that the domain-adapted images achieve superior scene text detection and recognition performance while applied to network training.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Model Adaptation: Historical Contrastive Learning for Unsupervised Domain Adaptation without Source DataJiaxing Huang, Dayan Guan, Aoran Xiao, Shijian LuNeurIPS 2021 · 被引用 301 次
- Diverse Image Inpainting with Bidirectional and Autoregressive TransformersYingchen Yu, Fangneng Zhan, Rongliang Wu, Jianxiong Pan 等ACM MM 2021 · 被引用 153 次
- Unsupervised Domain Adaptive 3D Detection with Multi-Level ConsistencyZhipeng Luo, Zhongang Cai, Changqing Zhou, Gongjie Zhang 等ICCV 2021 · 被引用 92 次
- RDA: Robust Domain Adaptation via Fourier Adversarial AttackingJiaxing Huang, Dayan Guan, Aoran Xiao, Shijian LuICCV 2021 · 被引用 85 次
- EMLight: Lighting Estimation via Spherical Distribution ApproximationFangneng Zhan, Changgong Zhang, Yingchen Yu, Yuan Chang 等AAAI 2021 · 被引用 73 次
相关 Paper
- Multimodal Structure-Consistent Image-to-Image TranslationChe-Tsung Lin, Yen-Yi Wu, Po-Hao Hsu, Shang-Hong LaiAAAI 2020 · 被引用 24 次
- Geometry-Aware Network for Domain Adaptive Semantic SegmentationYinghong Liao, Wending Zhou, Xu Yan, Zhen Li 等AAAI 2023 · 被引用 9 次
- Multi-Source Domain Adaptation for Visual Sentiment ClassificationChuang Lin, Sicheng Zhao, Lei Meng, Tat-Seng ChuaAAAI 2020 · 被引用 78 次
- Geometry Normalization Networks for Accurate Scene Text DetectionJiaqi Duan, Youjiang Xu, Zhanghui Kuang, Xiaoyu Yue 等ICCV 2019 · 被引用 33 次
- DRANet: Disentangling Representation and Adaptation Networks for Unsupervised Cross-Domain AdaptationSeunghun Lee, Sunghyun Cho, Sunghoon ImCVPR 2021
