Video Synthesis via Transform-Based Tensor Neural Network
Yimeng Zhang, Xiao-Yang Liu, Bo Wu, Anwar Walid
摘要
Video frame synthesis is an important task in computer vision and has drawn great interests in wide applications. However, existing neural network methods do not explicitly impose tensor low-rankness of videos to capture the spatiotemporal correlations in a high-dimensional space, while existing iterative algorithms require hand-crafted parameters and take relatively long running time. In this paper, we propose a novel multi-phase deep neural network Transform-Based Tensor-Net that exploits the low-rank structure of video data in a learned transform domain, which unfolds an Iterative Shrinkage-Thresholding Algorithm (ISTA) for tensor signal recovery. Our design is based on two observations: (i) both linear and nonlinear transforms can be implemented by a neural network layer, and (ii) the soft-thresholding operator corresponds to an activation function. Further, such an unfolding design is able to achieve nearly real-time at the cost of training time and enjoys an interpretable nature as a byproduct. Experimental results on the KTH and UCF-101 datasets show that compared with the state-of-the-art methods, i.e., DVF and Super SloMo, the proposed scheme improves Peak Signal-to-Noise Ratio (PSNR) of video interpolation and prediction by 4.13 dB and 4.26 dB, respectively.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- HLRTF: Hierarchical Low-Rank Tensor Factorization for Inverse Problems in Multi-Dimensional ImagingYi-Si Luo, Xile Zhao, Deyu Meng, Tai-Xiang JiangCVPR 2022 · 被引用 45 次
- Tensor FISTA-Net for Real-Time Snapshot Compressive ImagingXiaochen Han, Bo Wu, Zheng Shou, Xiao-Yang Liu 等AAAI 2020 · 被引用 46 次
- Deep Tensor ADMM-Net for Snapshot Compressive ImagingJiawei Ma, Xiao-Yang Liu, Zheng Shou, Xin YuanICCV 2019 · 被引用 218 次
- RSTT: Real-time Spatial Temporal Transformer for Space-Time Video Super-ResolutionZhicheng Geng, Luming Liang, Tianyu Ding, Ilya ZharkovCVPR 2022 · 被引用 103 次
- ST-MFNet: A Spatio-Temporal Multi-Flow Network for Frame InterpolationDuolikun Danier, Fan Zhang, David BullCVPR 2022 · 被引用 46 次
