CFFT-GAN: Cross-Domain Feature Fusion Transformer for Exemplar-Based Image Translation
Tianxiang Ma, Bingchuan Li, Wei Liu, Miao Hua, Jing Dong, Tieniu Tan
Abstract
Exemplar-based image translation refers to the task of generating images with the desired style, while conditioning on certain input image. Most of the current methods learn the correspondence between two input domains and lack the mining of information within the domain. In this paper, we propose a more general learning approach by considering two domain features as a whole and learning both inter-domain correspondence and intra-domain potential information interactions. Specifically, we propose a Cross-domain Feature Fusion Transformer (CFFT) to learn inter- and intra-domain feature fusion. Based on CFFT, the proposed CFFT-GAN works well on exemplar-based image translation. Moreover, CFFT-GAN is able to decouple and fuse features from multiple domains by cascading CFFT modules. We conduct rich quantitative and qualitative experiments on several image translation tasks, and the results demonstrate the superiority of our approach compared to state-of-the-art methods. Ablation studies show the importance of our proposed CFFT. Application experimental results reflect the potential of our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ddd396ea-eb22-42e9-9e81-ac1110248036Cited by top-tier papers1
Ask how each one uses itBuilds on18
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- ViViT: A Video Vision TransformerAnurag Arnab, Mostafa Dehghani, Georg Heigold, Chen Sun et al.ICCV 2021 · 2,947 citations
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine et al.NeurIPS 2020 · 2,345 citations
- Learning an animatable detailed 3D face model from in-the-wild imagesYao Feng, Haiwen Feng, Michael J. Black, Timo BolkartSIGGRAPH 2021 · 662 citations
- TransGAN: Two Pure Transformers Can Make One Strong GAN, and That Can Scale UpYifan Jiang, Shiyu Chang, Zhangyang WangNeurIPS 2021 · 515 citations
Related papers
- Cross-Domain Correspondence Learning for Exemplar-Based Image TranslationPan Zhang, Bo Zhang, Dong Chen, Lu Yuan et al.CVPR 2020
- Masked and Adaptive Transformer for Exemplar Based Image TranslationChang Jiang, Fei Gao, Biao Ma, Yuhao Lin et al.CVPR 2023
- Cross-Granularity Learning for Multi-Domain Image-to-Image TranslationHuiyuan Fu, Ting Yu, Xin Wang, Huadong MaACM MM 2020 · 4 citations
- Unbalanced Feature Transport for Exemplar-Based Image TranslationFangneng Zhan, Yingchen Yu, Kaiwen Cui, Gongjie Zhang et al.CVPR 2021
- MIDMs: Matching Interleaved Diffusion Models for Exemplar-Based Image TranslationJunyoung Seo, Gyuseong Lee, Seokju Cho, Jiyoung Lee et al.AAAI 2023 · 32 citations
