Audio-Assisted Face Video Restoration with Temporal and Identity Complementary Learning
Yuqin Cao, Yixuan Gao, Wei Sun, Xiaohong Liu, Yulun Zhang, Xiongkuo Min
Abstract
Face videos accompanied by audio have become integral to our daily lives, while they often suffer from complex degradations. Most face video restoration methods neglect the intrinsic correlations between the visual and audio features, especially in mouth regions. A few audio-aided face video restoration methods have been proposed, but they only focus on compression artifact removal. In this paper, we propose a General Audio-assisted face Video restoration Network (GAVN) to address various types of streaming video distortions via identity and temporal complementary learning. Specifically, GAVN first captures inter-frame temporal features in the low-resolution space to restore frames coarsely and save computational cost. Then, GAVN extracts intraframe identity features in the high-resolution space with the assistance of audio signals and face landmarks to restore more facial details. Finally, the reconstruction module integrates temporal features and identity features to generate high-quality face videos. Experimental results demonstrate that GAVN outperforms the existing state-of-the-art methods on face video compression artifact removal, deblurring, and super-resolution. Codes will be released upon publication.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 84cb4d38-dc40-4c01-9440-0793ddc42dc1Builds on10
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 522 citations
- A Differentiable Two-stage Alignment Scheme for Burst Image Reconstruction with Large ShiftShi Guo, Xi Yang, Jianqi Ma, Gaofeng Ren et al.CVPR 2022 · 14 citations
- Deep Multi-modality Soft-decoding of Very Low Bit-rate Face VideosYanhui Guo, Xi Zhang, Xiaolin WuACM MM 2020 · 9 citations
- TDAN: Temporally-Deformable Alignment Network for Video Super-ResolutionYapeng Tian, Yulun Zhang, Yun Fu, Chenliang XuCVPR 2020
- DR2: Diffusion-Based Robust Degradation Remover for Blind Face RestorationZhixin Wang, Ziying Zhang, Xiaoyun Zhang, Huangjie Zheng et al.CVPR 2023
Related papers
- Show and Polish: Reference-Guided Identity Preservation in Face Video RestorationWenkang Han, Wang Lin, Yiyun Zhou, Qi Liu et al.ACM MM 2025 · 1 citation
- Face Video Deblurring Using 3D Facial PriorsWenqi Ren, Jiaolong Yang, Senyou Deng, David P. Wipf et al.ICCV 2019 · 52 citations
- Transcoded Video Restoration by Temporal Spatial Auxiliary NetworkLi Xu, Gang He, Jinjia Zhou, Jie Lei et al.AAAI 2022 · 17 citations
- Exploring Correlations in Degraded Spatial Identity Features for Blind Face RestorationQian Ning, Fangfang Wu, Weisheng Dong, Xin Li et al.ACM MM 2023 · 1 citation
- A Simple Baseline for Video Restoration with Grouped Spatial-Temporal ShiftDasong Li, Xiaoyu Shi, Yi Zhang, Ka Chun Cheung et al.CVPR 2023
