TDAN: Temporally-Deformable Alignment Network for Video Super-Resolution
Yapeng Tian, Yulun Zhang, Yun Fu, Chenliang Xu
Abstract
Video super-resolution (VSR) aims to restore a photorealistic high-resolution (HR) video frame from both its corresponding low-resolution (LR) frame (reference frame) and multiple neighboring frames (supporting frames). Due to varying motion of cameras or objects, the reference frame and each support frame are not aligned. Therefore, temporal alignment is a challenging yet important problem for VSR. Previous VSR methods usually utilize optical flow between the reference frame and each supporting frame to wrap the supporting frame for temporal alignment. Therefore, the performance of these image-level wrapping-based models will highly depend on the prediction accuracy of optical flow, and inaccurate optical flow will lead to artifacts in the wrapped supporting frames, which also will be propagated into the reconstructed HR video frame. To overcome the limitation, in this paper, we propose a temporal deformable alignment network (TDAN) to adaptively align the reference frame and each supporting frame at the feature level without computing optical flow. The TDAN uses features from both the reference frame and each supporting frame to dynamically predict offsets of sampling convolution kernels. By using the corresponding kernels, TDAN transforms supporting frames to align with the reference frame. To predict the HR video frame, a reconstruction network taking aligned frames and the reference frame is utilized. Experimental results demonstrate the effectiveness of the proposed TDAN-based VSR model.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2c620519-3420-496a-bd64-5cdf8a809842Cited by top-tier papers117
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 522 citations
- Recurrent Video Restoration Transformer with Guided Deformable AttentionJingyun Liang, Yuchen Fan, Xiaoyu Xiang, Rakesh Ranjan et al.NeurIPS 2022 · 318 citations
- FaPN: Feature-aligned Pyramid Network for Dense Image PredictionShihua Huang, Zhichao Lu, Ran Cheng, Cheng HeICCV 2021 · 256 citations
- XVFI: eXtreme Video Frame InterpolationHyeonjun Sim, Jihyong Oh, Munchurl KimICCV 2021 · 207 citations
- Understanding Deformable Alignment in Video Super-ResolutionKelvin C. K. Chan, Xintao Wang, Ke Yu, Chao Dong et al.AAAI 2021 · 184 citations
Related papers
- Spatio-Temporal Filter Adaptive Network for Video DeblurringShangchen Zhou, Jiawei Zhang, Jinshan Pan, Wangmeng Zuo et al.ICCV 2019 · 225 citations
- Look Back and Forth: Video Super-Resolution with Explicit Temporal Difference ModelingTakashi Isobe, Xu Jia, Xin Tao, Changlin Li et al.CVPR 2022 · 57 citations
- Event-Enhanced Blurry Video Super-ResolutionDachun Kai, Yueyi Zhang, Jin Wang, Zeyu Xiao et al.AAAI 2025 · 7 citations
- Temporal Modulation Network for Controllable Space-Time Video Super-ResolutionGang Xu, Jun Xu, Zhen Li, Liang Wang et al.CVPR 2021
- Progressive Temporal Feature Alignment Network for Video InpaintingXueyan Zou, Linjie Yang, Ding Liu, Yong Jae LeeCVPR 2021
