Audio-Assisted Face Video Restoration with Temporal and Identity Complementary Learning
Yuqin Cao, Yixuan Gao, Wei Sun, Xiaohong Liu, Yulun Zhang, Xiongkuo Min
摘要
Face videos accompanied by audio have become integral to our daily lives, while they often suffer from complex degradations. Most face video restoration methods neglect the intrinsic correlations between the visual and audio features, especially in mouth regions. A few audio-aided face video restoration methods have been proposed, but they only focus on compression artifact removal. In this paper, we propose a General Audio-assisted face Video restoration Network (GAVN) to address various types of streaming video distortions via identity and temporal complementary learning. Specifically, GAVN first captures inter-frame temporal features in the low-resolution space to restore frames coarsely and save computational cost. Then, GAVN extracts intraframe identity features in the high-resolution space with the assistance of audio signals and face landmarks to restore more facial details. Finally, the reconstruction module integrates temporal features and identity features to generate high-quality face videos. Experimental results demonstrate that GAVN outperforms the existing state-of-the-art methods on face video compression artifact removal, deblurring, and super-resolution. Codes will be released upon publication.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper10
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 被引用 522 次
- A Differentiable Two-stage Alignment Scheme for Burst Image Reconstruction with Large ShiftShi Guo, Xi Yang, Jianqi Ma, Gaofeng Ren 等CVPR 2022 · 被引用 14 次
- Deep Multi-modality Soft-decoding of Very Low Bit-rate Face VideosYanhui Guo, Xi Zhang, Xiaolin WuACM MM 2020 · 被引用 9 次
- TDAN: Temporally-Deformable Alignment Network for Video Super-ResolutionYapeng Tian, Yulun Zhang, Yun Fu, Chenliang XuCVPR 2020
- DR2: Diffusion-Based Robust Degradation Remover for Blind Face RestorationZhixin Wang, Ziying Zhang, Xiaoyun Zhang, Huangjie Zheng 等CVPR 2023
相关 Paper
- Show and Polish: Reference-Guided Identity Preservation in Face Video RestorationWenkang Han, Wang Lin, Yiyun Zhou, Qi Liu 等ACM MM 2025 · 被引用 1 次
- Face Video Deblurring Using 3D Facial PriorsWenqi Ren, Jiaolong Yang, Senyou Deng, David P. Wipf 等ICCV 2019 · 被引用 52 次
- Transcoded Video Restoration by Temporal Spatial Auxiliary NetworkLi Xu, Gang He, Jinjia Zhou, Jie Lei 等AAAI 2022 · 被引用 17 次
- Exploring Correlations in Degraded Spatial Identity Features for Blind Face RestorationQian Ning, Fangfang Wu, Weisheng Dong, Xin Li 等ACM MM 2023 · 被引用 1 次
- A Simple Baseline for Video Restoration with Grouped Spatial-Temporal ShiftDasong Li, Xiaoyu Shi, Yi Zhang, Ka Chun Cheung 等CVPR 2023
