Unsupervised Video Interpolation Using Cycle Consistency
Fitsum A. Reda, Deqing Sun, Aysegul Dundar, Mohammad Shoeybi, Guilin Liu, Kevin J. Shih, Andrew Tao, Jan Kautz, Bryan Catanzaro
Abstract
Learning to synthesize high frame rate videos via interpolation requires large quantities of high frame rate training videos, which, however, are scarce, especially at high resolutions. Here, we propose unsupervised techniques to synthesize high frame rate videos directly from low frame rate videos using cycle consistency. For a triplet of consecutive frames, we optimize models to minimize the discrepancy between the center frame and its cycle reconstruction, obtained by interpolating back from interpolated intermediate frames. This simple unsupervised constraint alone achieves results comparable with supervision using the ground truth intermediate frames. We further introduce a pseudo supervised loss term that enforces the interpolated frames to be consistent with predictions of a pre-trained interpolation model. The pseudo supervised loss term, used together with cycle consistency, can effectively adapt a pre-trained model to a new target domain. With no additional data and in a completely unsupervised fashion, our techniques significantly improve pretrained models on new target domains, increasing PSNR values from 32.84dB to 33.05dB on the Slowflow and from 31.82dB to 32.53dB on the Sintel evaluation datasets. Code is available at https://github.com/NVIDIA/unsupervisedvideo-interpolation .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers30
- Asymmetric Bilateral Motion Estimation for Video Frame InterpolationJunheum Park, Chul Lee, Chang-Su KimICCV 2021 · 186 citations
- Video Frame Interpolation with TransformerLiying Lu, Ruizheng Wu, Huaijia Lin, Jiangbo Lu et al.CVPR 2022 · 128 citations
- VideoINR: Learning Video Implicit Neural Representation for Continuous Space-Time Super-ResolutionZeyuan Chen, Yinbo Chen, Jingwen Liu, Xingqian Xu et al.CVPR 2022 · 95 citations
- Many-to-many Splatting for Efficient Video Frame InterpolationPing Hu, Simon Niklaus, Stan Sclaroff, Kate SaenkoCVPR 2022 · 63 citations
- Vid-ODE: Continuous-Time Video Generation with Neural Ordinary Differential EquationSunghyun Park, Kangyeol Kim, Junsoo Lee, Jaegul Choo et al.AAAI 2021 · 62 citations
Related papers
- Learning Temporally and Semantically Consistent Unpaired Video-to-Video Translation through Pseudo-Supervision from Synthetic Optical FlowKaihong Wang, Kumar Akash, Teruhisa MisuAAAI 2022 · 16 citations
- Semi-Supervised Video Inpainting with Cycle Consistency ConstraintsZhiliang Wu, Hanyu Xuan, Changchang Sun, Weili Guan et al.CVPR 2023
- Task Agnostic Restoration of Natural Video DynamicsMuhammad Kashif Ali, Dongjin Kim, Tae Hyun KimICCV 2023
- DistractFlow: Improving Optical Flow Estimation via Realistic Distractions and Pseudo-LabelingJisoo Jeong, Hong Cai, Risheek Garrepalli, Fatih PorikliCVPR 2023
- Video Frame Interpolation without Temporal PriorsYoujian Zhang, Chaoyue Wang, Dacheng TaoNeurIPS 2020 · 38 citations
