Video Dynamics Prior: An Internal Learning Approach for Robust Video Enhancements
Gaurav Shrivastava, Ser Nam Lim, Abhinav Shrivastava
Abstract
In this paper, we present a novel robust framework for low-level vision tasks, including denoising, object removal, frame interpolation, and super-resolution, that does not require any external training data corpus. Our proposed approach directly learns the weights of neural modules by optimizing over the corrupted test sequence, leveraging the spatio-temporal coherence and internal statistics of videos. Furthermore, we introduce a novel spatial pyramid loss that leverages the property of spatio-temporal patch recurrence in a video across the different scales of the video. This loss enhances robustness to unstructured noise in both the spatial and temporal domains. This further results in our framework being highly robust to degradation in input frames and yields state-of-the-art results on downstream tasks such as denoising, object removal, and frame interpolation. To validate the effectiveness of our approach, we conduct qualitative and quantitative evaluations on standard video datasets such as DAVIS, UCF-101, and VIMEO90K-T. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on13
- SinGAN: Learning a Generative Model From a Single Natural ImageTamar Rott Shaham, Tali Dekel, Tomer MichaeliICCV 2019 · 933 citations
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 522 citations
- InGAN: Capturing and Retargeting the "DNA" of a Natural ImageAssaf Shocher, Shai Bagon, Phillip Isola, Michal IraniICCV 2019 · 146 citations
- Blind Video Temporal Consistency via Deep Video PriorChenyang Lei, Yazhou Xing, Qifeng ChenNeurIPS 2020 · 134 citations
- Unsupervised Deep Video DenoisingDev Yashpal Sheth, Sreyas Mohan, Joshua L. Vincent, Ramon Manzorro et al.ICCV 2021 · 78 citations
Related papers
- Instant Video Models: Universal Adapters for Stabilizing Image-Based NetworksMatthew Dutson, Nathan Labiosa, Yin Li, Mohit GuptaNeurIPS 2025
- Spatio-temporal Prompting Network for Robust Video Feature ExtractionGuanxiong Sun, Chi Wang, Zhaoyu Zhang, Jiankang Deng et al.ICCV 2023 · 11 citations
- Temporal Coherent Test Time Optimization for Robust Video ClassificationChenyu Yi, Siyuan Yang, Yufei Wang, Haoliang Li et al.ICLR 2023 · 3 citations
- Unsupervised Deep Video Denoising with Untrained NetworkHuan Zheng, Tongyao Pang, Hui JiAAAI 2023 · 14 citations
- A Simple Baseline for Video Restoration with Grouped Spatial-Temporal ShiftDasong Li, Xiaoyu Shi, Yi Zhang, Ka Chun Cheung et al.CVPR 2023
