Frequency Domain Image Translation: More Photo-realistic, Better Identity-preserving
Mu Cai, Hong Zhang, Huijuan Huang, Qichuan Geng, Yixuan Li, Gao Huang
摘要
Image-to-image translation has been revolutionized with GAN-based methods. However, existing methods lack the ability to preserve the identity of the source domain. As a result, synthesized images can often over-adapt to the reference domain, losing important structural characteristics and suffering from suboptimal visual quality. To solve these challenges, we propose a novel frequency domain image translation (FDIT) framework, exploiting frequency information for enhancing the image generation process. Our key idea is to decompose the image into low-frequency and high-frequency components, where the high-frequency feature captures object structure akin to the identity. Our training objective facilitates the preservation of frequency information in both pixel space and Fourier spectral space. We broadly evaluate FDIT across five large-scale datasets and multiple tasks including image translation and GAN inversion. Extensive experiments and ablations show that FDIT effectively preserves the identity of the source image, and produces photo-realistic images. FDIT establishes state-of-the-art performance, reducing the average FID score by 5.6% compared to the previous best method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper23
- Focal Frequency Loss for Image Reconstruction and SynthesisLiming Jiang, Bo Dai, Wayne Wu, Chen Change LoyICCV 2021 · 被引用 422 次
- Spectral Unsupervised Domain Adaptation for Visual RecognitionJingyi Zhang, Jiaxing Huang, Zichen Tian, Shijian LuCVPR 2022 · 被引用 72 次
- AesFA: An Aesthetic Feature-Aware Arbitrary Neural Style TransferJoonwoo Kwon, Sooyoung Kim, Yuewei Lin, Shinjae Yoo 等AAAI 2024 · 被引用 32 次
- Painterly Image Harmonization in Dual DomainsJunyan Cao, Yan Hong, Li NiuAAAI 2023 · 被引用 27 次
- Frequency-Controlled Diffusion Model for Versatile Text-Guided Image-to-Image TranslationXiang Gao, Zhengbo Xu, Junhan Zhao, Jiaying LiuAAAI 2024 · 被引用 23 次
它引用的顶会 Paper17
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 被引用 1,195 次
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 被引用 1,049 次
- Drop an Octave: Reducing Spatial Redundancy in Convolutional Neural Networks With Octave ConvolutionYunpeng Chen, Haoqi Fan, Bing Xu, Zhicheng Yan 等ICCV 2019 · 被引用 665 次
- Photorealistic Style Transfer via Wavelet TransformsJaejun Yoo, Youngjung Uh, Sanghyuk Chun, Byeongkyu Kang 等ICCV 2019 · 被引用 412 次
- Swapping Autoencoder for Deep Image ManipulationTaesung Park, Jun-Yan Zhu, Oliver Wang, Jingwan Lu 等NeurIPS 2020 · 被引用 376 次
相关 Paper
- SEIT: Structural Enhancement for Unsupervised Image Translation in Frequency DomainZhifeng Zhu, Yaochen Li, Yifan Li, Jinhuo Yang 等AAAI 2024 · 被引用 5 次
- Frequency-Guided Diffusion for Training-Free Text-Driven Image TranslationZheng Gao, Jifei Song, Zhensong Zhang, Jiankang Deng 等ICCV 2025 · 被引用 1 次
- FSDR: Frequency Space Domain Randomization for Domain GeneralizationJiaxing Huang, Dayan Guan, Aoran Xiao, Shijian LuCVPR 2021
- Spectrum Translation for Refinement of Image Generation (STIG) Based on Contrastive Learning and Spectral Filter ProfileSeokjun Lee, Seung-Won Jung, Hyunseok SeoAAAI 2024 · 被引用 8 次
- Wavelet Knowledge Distillation: Towards Efficient Image-to-Image TranslationLinfeng Zhang, Xin Chen, Xiaobing Tu, Pengfei Wan 等CVPR 2022 · 被引用 105 次
