Optimizing Video Prediction via Video Frame Interpolation
Yue Wu, Qiang Wen, Qifeng Chen
Abstract
Video prediction is an extrapolation task that predicts future frames given past frames, and video frame interpolation is an interpolation task that estimates intermediate frames between two frames. We have witnessed the tremendous advancement of video frame interpolation, but the general video prediction in the wild is still an open question. Inspired by the photo-realistic results of video frame interpolation, we present a new optimization framework for video prediction via video frame interpolation, in which we solve an extrapolation problem based on an interpolation model. Our video prediction framework is based on optimization with a pretrained differentiable video frame interpolation module without the need for a training dataset, and thus there is no domain gap issue between training and test data. Also, our approach does not need any additional information such as semantic or instance maps, which makes our framework applicable to any video. Extensive experiments on the Cityscapes, KITTI, DAVIS, Middlebury, and Vimeo90K datasets show that our video prediction results are robust in general scenarios, and our approach outperforms other video prediction methods that require a large amount of training data or extra semantic information.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1dacfb77-69aa-4346-96d2-89e7962746b1Cited by top-tier papers14
- MMVP: Motion-Matrix-based Video PredictionYiqi Zhong, Luming Liang, Ilya Zharkov, Ulrich NeumannICCV 2023 · 39 citations
- PastNet: Introducing Physical Inductive Biases for Spatio-temporal Video PredictionHao Wu, Fan Xu, Chong Chen, Xian-Sheng Hua et al.ACM MM 2024 · 36 citations
- STDiff: Spatio-Temporal Diffusion for Continuous Stochastic Video PredictionXi Ye, Guillaume-Alexandre BilodeauAAAI 2024 · 20 citations
- Vid-SME: Membership Inference Attacks against Large Video Understanding ModelsQi Li, Runpeng Yu, Xinchao WangNeurIPS 2025 · 19 citations
- Advection Augmented Convolutional Neural NetworksNiloufar Zakariaei, Siddharth Rout, Eldad Haber, Moshe EliasofNeurIPS 2024 · 7 citations
Builds on10
- SinGAN: Learning a Generative Model From a Single Natural ImageTamar Rott Shaham, Tali Dekel, Tomer MichaeliICCV 2019 · 933 citations
- XVFI: eXtreme Video Frame InterpolationHyeonjun Sim, Jihyong Oh, Munchurl KimICCV 2021 · 207 citations
- Asymmetric Bilateral Motion Estimation for Video Frame InterpolationJunheum Park, Chul Lee, Chang-Su KimICCV 2021 · 186 citations
- Blind Video Temporal Consistency via Deep Video PriorChenyang Lei, Yazhou Xing, Qifeng ChenNeurIPS 2020 · 134 citations
- Disentangling Propagation and Generation for Video PredictionHang Gao, Huazhe Xu, Qi-Zhi Cai, Ruth Wang et al.ICCV 2019 · 90 citations
Related papers
- Future Video Synthesis With Object Motion PredictionYue Wu, Rongrong Gao, Jaesik Park, Qifeng ChenCVPR 2020
- Repurposing Pre-trained Video Diffusion Models for Event-based Video InterpolationJingxi Chen, Brandon Y. Feng, Haoming Cai, Tianfu Wang et al.CVPR 2025
- MCVD - Masked Conditional Video Diffusion for Prediction, Generation, and InterpolationVikram Voleti, Alexia Jolicoeur-Martineau, Chris PalNeurIPS 2022 · 434 citations
- ExtDM: Distribution Extrapolation Diffusion Model for Video PredictionZhicheng Zhang, Junyao Hu, Wentao Cheng, Danda Pani Paudel et al.CVPR 2024 · 24 citations
- Long-term Video Frame Interpolation via Feature PropagationDawit Mureja Argaw, In So KweonCVPR 2022 · 12 citations
