Transcoded Video Restoration by Temporal Spatial Auxiliary Network
Li Xu, Gang He, Jinjia Zhou, Jie Lei, Weiying Xie, Yunsong Li, Yu-Wing Tai
Abstract
In most video platforms, such as Youtube, Kwai, and TikTok, the played videos usually have undergone multiple video encodings such as hardware encoding by recording devices, software encoding by video editing apps, and single/multiple video transcoding by video application servers. Previous works in compressed video restoration typically assume the compression artifacts are caused by one-time encoding. Thus, the derived solution usually does not work very well in practice. In this paper, we propose a new method, temporal spatial auxiliary network (TSAN), for transcoded video restoration. Our method considers the unique traits between video encoding and transcoding, and we consider the initial shallow encoded videos as the intermediate labels to assist the network to conduct self-supervised attention training. In addition, we employ adjacent multi-frame information and propose the temporal deformable alignment and pyramidal spatial fusion for transcoded video restoration. The experimental results demonstrate that the performance of the proposed method is superior to that of the previous techniques. The code is available at https://github.com/icecherylXuli/TSAN.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- CPGA: Coding Priors-Guided Aggregation Network for Compressed Video Quality EnhancementQiang Zhu, Jinhua Hao, Yukang Ding, Yu Liu et al.CVPR 2024 · 13 citations
- Compression-Aware Video Super-ResolutionYingwei Wang, Takashi Isobe, Xu Jia, Xin Tao et al.CVPR 2023
- Perceptual Video Compression with Neural WrappingMuhammad Umar Karim Khan, Aaron Chadha, Mohammad Ashraful Anam, Yiannis AndreopoulosCVPR 2025
Builds on4
- Understanding Deformable Alignment in Video Super-ResolutionKelvin C. K. Chan, Xintao Wang, Ke Yu, Chao Dong et al.AAAI 2021 · 184 citations
- Spatio-Temporal Deformable Convolution for Compressed Video Quality EnhancementJianing Deng, Li Wang, Shiliang Pu, Cheng ZhuoAAAI 2020 · 168 citations
- EfficientDeRain: Learning Pixel-wise Dilation Filtering for High-Efficiency Single-Image DerainingQing Guo, Jingyang Sun, Felix Juefei-Xu, Lei Ma et al.AAAI 2021 · 120 citations
- TDAN: Temporally-Deformable Alignment Network for Video Super-ResolutionYapeng Tian, Yulun Zhang, Yun Fu, Chenliang XuCVPR 2020
Related papers
- Audio-Assisted Face Video Restoration with Temporal and Identity Complementary LearningYuqin Cao, Yixuan Gao, Wei Sun, Xiaohong Liu et al.AAAI 2026
- Deep Multi-modality Soft-decoding of Very Low Bit-rate Face VideosYanhui Guo, Xi Zhang, Xiaolin WuACM MM 2020 · 9 citations
- Recursive Fusion and Deformable Spatiotemporal Attention for Video Compression Artifact ReductionMinyi Zhao, Yi Xu, Shuigeng ZhouACM MM 2021 · 61 citations
- Video Compression Artifact Reduction by Fusing Motion Compensation and Global Context in a Swin-CNN Based Parallel ArchitectureXinjian Zhang, Su Yang, Wuyang Luo, Longwen Gao et al.AAAI 2023 · 15 citations
- A Simple Baseline for Video Restoration with Grouped Spatial-Temporal ShiftDasong Li, Xiaoyu Shi, Yi Zhang, Ka Chun Cheung et al.CVPR 2023
