RefSTAR: Blind Face Image Restoration with Reference Selection, Transfer, and Reconstruction
Zhicun Yin, Junjie Chen, Ming Liu, Zhixin Wang, Fan Li, Renjing Pei, Xiaoming Li, Rynson W. H. Lau, Wangmeng Zuo
Abstract
Blind facial image restoration is highly challenging due to unknown complex degradations and the sensitivity of humans to faces. Although existing methods introduce auxiliary information from generative priors or high-quality reference images, they still struggle with identity preservation problems, mainly due to improper feature introduction on detailed textures. In this paper, we focus on effectively incorporating appropriate features from high-quality reference images, presenting a novel blind facial image restoration method that considers reference selection, transfer, and reconstruction (RefSTAR). In terms of selection, we construct a reference selection (RefSel) module. For training the RefSel module, we construct a RefSel-HQ dataset through a mask generation pipeline, which contains annotating masks for 10,000 ground truth-reference pairs. As for the transfer, due to the trivial solution in vanilla crossattention operations, a feature fusion paradigm is designed to force the features from the reference to be integrated. Finally, we propose a reference image reconstruction mechanism that further ensures the presence of reference image features in the output image. The cycle consistency loss is also redesigned in conjunction with the mask. Extensive experiments on various backbone models demonstrate superior performance, showing better identity preservation ability and reference feature transfer quality. Source code, dataset, and pre-trained models are available at https: //github.com/yinzhicun/RefSTAR .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on17
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- MUSIQ: Multi-scale Image Quality TransformerJunjie Ke, Qifei Wang, Yilin Wang, Peyman Milanfar et al.ICCV 2021 · 1,325 citations
- Towards Robust Blind Face Restoration with Codebook Lookup TransformerShangchen Zhou, Kelvin C. K. Chan, Chongyi Li, Chen Change LoyNeurIPS 2022 · 431 citations
- One-Step Effective Diffusion Network for Real-World Image Super-ResolutionRongyuan Wu, Lingchen Sun, Zhiyuan Ma, Lei ZhangNeurIPS 2024 · 319 citations
- HiFaceGAN: Face Renovation via Collaborative Suppression and ReplenishmentLingbo Yang, Shanshe Wang, Siwei Ma, Wen Gao et al.ACM MM 2020 · 136 citations
Related papers
- FaceMe: Robust Blind Face Restoration with Personal IdentificationSiyu Liu, Zheng-Peng Duan, Jia Ouyang, Jiayi Fu et al.AAAI 2025 · 18 citations
- ReF-LDM: A Latent Diffusion Model for Reference-based Face Image RestorationChi-Wei Hsiao, Yu-Lun Liu, Cheng-Kun Yang, Sheng-Po Kuo et al.NeurIPS 2024 · 20 citations
- Blind Face Restoration via Integrating Face Shape and Generative PriorsFeida Zhu, Junwei Zhu, Wenqing Chu, Xinyi Zhang et al.CVPR 2022 · 45 citations
- AuthFace: Towards Authentic Blind Face Restoration with Face-oriented Generative Diffusion PriorGuoqiang Liang, Qingnan Fan, Bingtao Fu, Jinwei Chen et al.ACM MM 2025 · 3 citations
- LD-BFR: Vector-Quantization-Based Face Restoration Model with Latent Diffusion EnhancementYuzhen Du, Teng Hu, Ran Yi, Lizhuang MaACM MM 2024 · 3 citations
