Recursive Fusion and Deformable Spatiotemporal Attention for Video Compression Artifact Reduction
Minyi Zhao, Yi Xu, Shuigeng Zhou
Abstract
A number of deep learning based algorithms have been proposed to recover high-quality videos from low-quality compressed ones. Among them, some restore the missing details of each frame via exploring the spatiotemporal information of neighboring frames. However, these methods usually suffer from a narrow temporal scope, thus may miss some useful details from some frames outside the neighboring ones. In this paper, to boost artifact removal, on the one hand, we propose a Recursive Fusion (RF) module to model the temporal dependency within a long temporal range. Specifically, RF utilizes both the current reference frames and the preceding hidden state to conduct better spatiotemporal compensation. On the other hand, we design an efficient and effective Deformable Spatiotemporal Attention (DSTA) module such that the model can pay more effort on restoring the artifact-rich areas like the boundary area of a moving object. Extensive experiments show that our method outperforms the existing ones on the MFQE 2.0 dataset in terms of both fidelity and perceptual effect. Code is available at https://github.com/zhaominyiz/RFDA-PyTorch.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 52a851ba-8d5e-4bf7-b9ca-bcf18e23c10eCited by top-tier papers7
- Learning Truncated Causal History Model for Video RestorationAmirhosein Ghasemabadi, Muhammad Kamran Janjua, Mohammad Salameh, Di NiuNeurIPS 2024 · 28 citations
- Video Compression Artifact Reduction by Fusing Motion Compensation and Global Context in a Swin-CNN Based Parallel ArchitectureXinjian Zhang, Su Yang, Wuyang Luo, Longwen Gao et al.AAAI 2023 · 15 citations
- CPGA: Coding Priors-Guided Aggregation Network for Compressed Video Quality EnhancementQiang Zhu, Jinhua Hao, Yukang Ding, Yu Liu et al.CVPR 2024 · 13 citations
- Keyword-Based Diverse Image Retrieval by Semantics-aware Contrastive Learning and TransformerMinyi Zhao, Jinpeng Wang, Dongliang Liao, Yiru Wang et al.SIGIR 2023 · 3 citations
- Motion Information Propagation for Neural Video CompressionLinfeng Qi, Jiahao Li, Bin Li, Houqiang Li et al.CVPR 2023
Builds on5
- Understanding Deformable Alignment in Video Super-ResolutionKelvin C. K. Chan, Xintao Wang, Ke Yu, Chao Dong et al.AAAI 2021 · 184 citations
- Spatio-Temporal Deformable Convolution for Compressed Video Quality EnhancementJianing Deng, Li Wang, Shiliang Pu, Cheng ZhuoAAAI 2020 · 168 citations
- Non-Local ConvLSTM for Video Compression Artifact ReductionYi Xu, Longwen Gao, Kai Tian, Shuigeng Zhou et al.ICCV 2019 · 70 citations
- GIF Thumbnails: Attract More Clicks to Your VideosYi Xu, Fan Bai, Yingxuan Shi, Qiuyu Chen et al.AAAI 2021 · 11 citations
- Residual Feature Aggregation Network for Image Super-ResolutionJie Liu, Wenjie Zhang, Yuting Tang, Jie Tang et al.CVPR 2020
Related papers
- Recurrent Video Restoration Transformer with Guided Deformable AttentionJingyun Liang, Yuchen Fan, Xiaoyu Xiang, Rakesh Ranjan et al.NeurIPS 2022 · 318 citations
- RivuletMLP: An MLP-based Architecture for Efficient Compressed Video Quality EnhancementGang He, Weiran Wang, Guancheng Quan, Shihao Wang et al.CVPR 2025
- Transcoded Video Restoration by Temporal Spatial Auxiliary NetworkLi Xu, Gang He, Jinjia Zhou, Jie Lei et al.AAAI 2022 · 17 citations
- Flow-Guided Sparse Transformer for Video DeblurringJing Lin, Yuanhao Cai, Xiaowan Hu, Haoqian Wang et al.ICML 2022 · 82 citations
- Gated Spatio-Temporal Attention-Guided Video DeblurringMaitreya Suin, A. N. RajagopalanCVPR 2021
