Video Synthesis via Transform-Based Tensor Neural Network
Yimeng Zhang, Xiao-Yang Liu, Bo Wu, Anwar Walid
Abstract
Video frame synthesis is an important task in computer vision and has drawn great interests in wide applications. However, existing neural network methods do not explicitly impose tensor low-rankness of videos to capture the spatiotemporal correlations in a high-dimensional space, while existing iterative algorithms require hand-crafted parameters and take relatively long running time. In this paper, we propose a novel multi-phase deep neural network Transform-Based Tensor-Net that exploits the low-rank structure of video data in a learned transform domain, which unfolds an Iterative Shrinkage-Thresholding Algorithm (ISTA) for tensor signal recovery. Our design is based on two observations: (i) both linear and nonlinear transforms can be implemented by a neural network layer, and (ii) the soft-thresholding operator corresponds to an activation function. Further, such an unfolding design is able to achieve nearly real-time at the cost of training time and enjoys an interpretable nature as a byproduct. Experimental results on the KTH and UCF-101 datasets show that compared with the state-of-the-art methods, i.e., DVF and Super SloMo, the proposed scheme improves Peak Signal-to-Noise Ratio (PSNR) of video interpolation and prediction by 4.13 dB and 4.26 dB, respectively.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get d3b78583-39ad-438e-99fb-047fd0ae9e11Cited by top-tier papers1
Ask how each one uses itRelated papers
- HLRTF: Hierarchical Low-Rank Tensor Factorization for Inverse Problems in Multi-Dimensional ImagingYi-Si Luo, Xile Zhao, Deyu Meng, Tai-Xiang JiangCVPR 2022 · 45 citations
- Tensor FISTA-Net for Real-Time Snapshot Compressive ImagingXiaochen Han, Bo Wu, Zheng Shou, Xiao-Yang Liu et al.AAAI 2020 · 46 citations
- Deep Tensor ADMM-Net for Snapshot Compressive ImagingJiawei Ma, Xiao-Yang Liu, Zheng Shou, Xin YuanICCV 2019 · 218 citations
- RSTT: Real-time Spatial Temporal Transformer for Space-Time Video Super-ResolutionZhicheng Geng, Luming Liang, Tianyu Ding, Ilya ZharkovCVPR 2022 · 103 citations
- ST-MFNet: A Spatio-Temporal Multi-Flow Network for Frame InterpolationDuolikun Danier, Fan Zhang, David BullCVPR 2022 · 46 citations
