Spatio-Temporal Deformable Convolution for Compressed Video Quality Enhancement
Jianing Deng, Li Wang, Shiliang Pu, Cheng Zhuo
Abstract
Recent years have witnessed remarkable success of deep learning methods in quality enhancement for compressed video. To better explore temporal information, existing methods usually estimate optical flow for temporal motion compensation. However, since compressed video could be seriously distorted by various compression artifacts, the estimated optical flow tends to be inaccurate and unreliable, thereby resulting in ineffective quality enhancement. In addition, optical flow estimation for consecutive frames is generally conducted in a pairwise manner, which is computational expensive and inefficient. In this paper, we propose a fast yet effective method for compressed video quality enhancement by incorporating a novel Spatio-Temporal Deformable Fusion (STDF) scheme to aggregate temporal information. Specifically, the proposed STDF takes a target frame along with its neighboring reference frames as input to jointly predict an offset field to deform the spatio-temporal sampling positions of convolution. As a result, complementary information from both target and reference frames can be fused within a single Spatio-Temporal Deformable Convolution (STDC) operation. Extensive experiments show that our method achieves the state-of-the-art performance of compressed video quality enhancement in terms of both accuracy and efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0faf4fb4-8ad1-459b-9f23-e41b413c450cCited by top-tier papers18
- Recursive Fusion and Deformable Spatiotemporal Attention for Video Compression Artifact ReductionMinyi Zhao, Yi Xu, Shuigeng ZhouACM MM 2021 · 61 citations
- COMISR: Compression-Informed Video Super-ResolutionYinxiao Li, Pengchong Jin, Feng Yang, Ce Liu et al.ICCV 2021 · 55 citations
- Learning Truncated Causal History Model for Video RestorationAmirhosein Ghasemabadi, Muhammad Kamran Janjua, Mohammad Salameh, Di NiuNeurIPS 2024 · 28 citations
- Unsupervised Flow-Aligned Sequence-to-Sequence Learning for Video RestorationJing Lin, Xiaowan Hu, Yuanhao Cai, Haoqian Wang et al.ICML 2022 · 27 citations
- AverNet: All-in-one Video Restoration for Time-varying Unknown DegradationsHaiyu Zhao, Lei Tian, Xinyan Xiao, Peng Hu et al.NeurIPS 2024 · 19 citations
Builds on1
Related papers
- Progressive Fusion Video Super-Resolution Network via Exploiting Non-Local Spatio-Temporal CorrelationsPeng Yi, Zhongyuan Wang, Kui Jiang, Junjun Jiang et al.ICCV 2019 · 309 citations
- Structure-Preserving Motion Estimation for Learned Video CompressionHan Gao, Jinzhong Cui, Mao Ye, Shuai Li et al.ACM MM 2022 · 16 citations
- Video Compression Artifact Reduction by Fusing Motion Compensation and Global Context in a Swin-CNN Based Parallel ArchitectureXinjian Zhang, Su Yang, Wuyang Luo, Longwen Gao et al.AAAI 2023 · 15 citations
- Multi-Frame Deformable Look-Up Table for Compressed Video Quality EnhancementGang He, Guancheng Quan, Chang Wu, Shihao Wang et al.AAAI 2025 · 4 citations
- RivuletMLP: An MLP-based Architecture for Efficient Compressed Video Quality EnhancementGang He, Weiran Wang, Guancheng Quan, Shihao Wang et al.CVPR 2025
