Video Super-Resolution using Multi-scale Pyramid 3D Convolutional Networks
Jianping Luo, Shaofei Huang, Yuan Yuan
Abstract
Video super-resolution (SR) aims at generating high-resolution (HR) frames from consecutive low-resolution (LR) frames. The challenge is how to make use of temporal coherence among neighbouring LR frames. Most previous works use motion estimation and compensation based models. However, their performance relies heavily on motion estimation accuracy. In this paper, we propose a multi-scale pyramid 3D convolutional (MP3D) network for video SR, where 3D convolution can explore temporal correlation directly without explicit motion compensation. Specifically, we first apply 3D convolution into a pyramid subnet to extractmulti-scale spatial and temporal features simultaneously from the LR frames, such that it can handle various sizes of motions. We then feed the fused feature maps into an SR reconstruction subnet, where a 3D sub-pixel convolution layer is used for up-sampling. Finally, we append a detail refinement subnet based on the encoder-decoder structure to further enhance texture details of the reconstructed HR frames. Extensive experiments on benchmark datasets and real-world cases show that the proposed MP3D model outperforms state-of-the-art video SR methods in terms of PSNR/SSIM values, visual quality and temporal consistency, respectively.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 3392ea6f-1ef8-48cc-8f7b-0fd5022d83aaCited by top-tier papers5
- Dense Deep Unfolding Network with 3D-CNN Prior for Snapshot Compressive ImagingZhuoyuan Wu, Jian Zhang, Chong MouICCV 2021 · 75 citations
- Multi-Frequency Representation Enhancement with Privilege Information for Video Super-ResolutionFei Li, Linfeng Zhang, Zikun Liu, Juan Lei et al.ICCV 2023 · 24 citations
- MambaSCI: Efficient Mamba-UNet for Quad-Bayer Patterned Video Snapshot Compressive ImagingZhenghao Pan, Haijin Zeng, Jiezhang Cao, Yongyong Chen et al.NeurIPS 2024 · 12 citations
- VSRM: A Robust Mamba-Based Framework for Video Super-ResolutionDinh Phu Tran, Dao Duy Hung, Daeyoung KimICCV 2025 · 4 citations
- EfficientSCI: Densely Connected Network with Space-time Factorization for Large-scale Video Snapshot Compressive ImagingLishun Wang, Miao Cao, Xin YuanCVPR 2023
Related papers
- Progressive Fusion Video Super-Resolution Network via Exploiting Non-Local Spatio-Temporal CorrelationsPeng Yi, Zhongyuan Wang, Kui Jiang, Junjun Jiang et al.ICCV 2019 · 309 citations
- Large Motion Video Super-Resolution with Dual Subnet and Multi-Stage Communicated UpsamplingHongying Liu, Peng Zhao, Zhubo Ruan, Fanhua Shang et al.AAAI 2021 · 28 citations
- Video Face Super-Resolution with Motion-Adaptive Feedback CellJingwei Xin, Nannan Wang, Jie Li, Xinbo Gao et al.AAAI 2020 · 14 citations
- How Video Super-Resolution and Frame Interpolation Mutually BenefitChengcheng Zhou, Zongqing Lu, Linge Li, Qiangyu Yan et al.ACM MM 2021 · 12 citations
- TDAN: Temporally-Deformable Alignment Network for Video Super-ResolutionYapeng Tian, Yulun Zhang, Yun Fu, Chenliang XuCVPR 2020
