Super-Resolving Cross-Domain Face Miniatures by Peeking at One-Shot Exemplar
Peike Li, Xin Yu, Yi Yang
Abstract
Conventional face super-resolution methods usually assume testing low-resolution (LR) images lie in the same domain as the training ones. Due to different lighting conditions and imaging hardware, domain gaps between training and testing images inevitably occur in many real-world scenarios. Neglecting those domain gaps would lead to inferior face super-resolution (FSR) performance. However, how to transfer a trained FSR model to a target domain efficiently and effectively has not been investigated. To tackle this problem, we develop a Domain-Aware Pyramid-based Face Super-Resolution network, named DAP-FSR network. Our DAP-FSR makes the first attempt to super-resolve LR faces from a target domain by exploiting only a pair of high-resolution (HR) and LR exemplar in the target domain. To be specific, our DAP-FSR firstly employs its encoder to extract the multi-scale latent representations of the input LR face. Considering only one target domain example is available, we propose to augment the target domain data by mixing the latent representations of the target domain face and source domain ones, and then feed the mixed representations to the decoder of our DAP-FSR. The decoder will generate new face images resembling the target domain image style. The generated HR faces in turn are used to optimize our decoder to reduce the domain gap. By iteratively updating the latent representations and our decoder, our DAP-FSR will be adapted to the target domain, thus achieving authentic and high-quality upsampled HR faces. Extensive experiments on three benchmarks validate the effectiveness and superior performance of our DAP-FSR compared to the state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- One-Shot Talking Face Generation from Single-Speaker Audio-Visual Correlation LearningSuzhen Wang, Lincheng Li, Yu Ding, Xin YuAAAI 2022 · 142 citations
- A Multi-Mode Modulator for Multi-Domain Few-Shot ClassificationYanbin Liu, Juho Lee, Linchao Zhu, Ling Chen et al.ICCV 2021 · 43 citations
Builds on9
- Few-Shot Unsupervised Image-to-Image TranslationMing-Yu Liu, Xun Huang, Arun Mallya, Tero Karras et al.ICCV 2019 · 668 citations
- SC-FEGAN: Face Editing Generative Adversarial Network With User's Sketch and ColorYoungjoo Jo, Jongyoul ParkICCV 2019 · 325 citations
- Adversarial Style Mining for One-Shot Unsupervised Domain AdaptationYawei Luo, Ping Liu, Tao Guan, Junqing Yu et al.NeurIPS 2020 · 129 citations
- Write-a-speaker: Text-based Emotional and Rhythmic Talking-head GenerationLincheng Li, Suzhen Wang, Zhimeng Zhang, Yu Ding et al.AAAI 2021 · 88 citations
- Attract or Distract: Exploit the Margin of Open SetQianyu Feng, Guoliang Kang, Hehe Fan, Yi YangICCV 2019 · 62 citations
Related papers
- Unsupervised Real-World Image Super Resolution via Domain-Distance Aware TrainingYunxuan Wei, Shuhang Gu, Yawei Li, Radu Timofte et al.CVPR 2021
- Unsupervised Real-World Super-Resolution: A Domain Adaptation PerspectiveWei Wang, Haochen Zhang, Zehuan Yuan, Changhu WangICCV 2021 · 67 citations
- IODA: Instance-Guided One-shot Domain Adaptation for Super-ResolutionZaizuo Tang, Yu-Bin YangNeurIPS 2024 · 3 citations
- Frequency Consistent Adaptation for Real World Super ResolutionXiaozhong Ji, Guangpin Tao, Yun Cao, Ying Tai et al.AAAI 2021 · 12 citations
- MASA-SR: Matching Acceleration and Spatial Adaptation for Reference-Based Image Super-ResolutionLiying Lu, Wenbo Li, Xin Tao, Jiangbo Lu et al.CVPR 2021
