SVFI: Spiking-Based Video Frame Interpolation for High-Speed Motion
Lujie Xia, Jing Zhao, Ruiqin Xiong, Tiejun Huang
Abstract
Occlusion and motion blur make it challenging to interpolate video frame, since estimating complex motions between two frames is hard and unreliable, especially in highly dynamic scenes. This paper aims to address these issues by exploiting spike stream as auxiliary visual information between frames to synthesize target frames. Instead of estimating motions by optical flow from RGB frames, we present a new dual-modal pipeline adopting both RGB frames and the corresponding spike stream as inputs (SVFI). It extracts the scene structure and objects' outline feature maps of the target frames from spike stream. Those feature maps are fused with the color and texture feature maps extracted from RGB frames to synthesize target frames. Benefited by the spike stream that contains consecutive information between two frames, SVFI can directly extract the information in occlusion and motion blur areas of target frames from spike stream, thus it is more robust than previous optical flow-based methods. Experiments show SVFI outperforms the SOTA methods on wide variety of datasets. For instance, in 7 and 15 frame skip evaluations, it shows up to 5.58 dB and 6.56 dB improvements in terms of PSNR over the corresponding second best methods BMBC and DAIN. SVFI also shows visually impressive performance in real-world scenes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Optical Flow for Spike Camera with Hierarchical Spatial-Temporal Spike FusionRui Zhao, Ruiqin Xiong, Jian Zhang, Xinfeng Zhang et al.AAAI 2024 · 23 citations
- Joint Demosaicing and Denoising for Spike CameraYanchen Dong, Ruiqin Xiong, Jing Zhao, Jian Zhang et al.AAAI 2024 · 18 citations
- Spike-guided Motion Deblurring with Unknown Modal Spatiotemporal AlignmentJiyuan Zhang, Shiyan Chen, Yajing Zheng, Zhaofei Yu et al.CVPR 2024 · 3 citations
- High Dynamic Range Imaging with Time-Encoding Spike CameraZhenkun Zhu, Ruiqin Xiong, Jiyu Xie, Yuanlin Wang et al.NeurIPS 2025
- Boosting Spike Camera Image Reconstruction from a Perspective of Dealing with Spike FluctuationsRui Zhao, Ruiqin Xiong, Jing Zhao, Jian Zhang et al.CVPR 2024
Builds on7
- Channel Attention Is All You Need for Video Frame InterpolationMyungsub Choi, Heewon Kim, Bohyung Han, Ning Xu et al.AAAI 2020 · 362 citations
- XVFI: eXtreme Video Frame InterpolationHyeonjun Sim, Jihyong Oh, Munchurl KimICCV 2021 · 207 citations
- Unsupervised Video Interpolation Using Cycle ConsistencyFitsum A. Reda, Deqing Sun, Aysegul Dundar, Mohammad Shoeybi et al.ICCV 2019 · 93 citations
- Extending Neural P-frame Codecs for B-frame CodingReza Pourreza, Taco CohenICCV 2021 · 52 citations
- Time Lens: Event-Based Video Frame InterpolationStepan Tulyakov, Daniel Gehrig, Stamatios Georgoulis, Julius Erbach et al.CVPR 2021
Related papers
- Enhancing Motion Deblurring in High-Speed Scenes with Spike StreamsShiyan Chen, Jiyuan Zhang, Yajing Zheng, Tiejun Huang et al.NeurIPS 2023 · 21 citations
- Optical Flow Estimation for Spiking CameraLiwen Hu, Rui Zhao, Ziluo Ding, Lei Ma et al.CVPR 2022 · 48 citations
- Learning Optical Flow from Continuous Spike StreamsRui Zhao, Ruiqin Xiong, Jing Zhao, Zhaofei Yu et al.NeurIPS 2022 · 49 citations
- Sparse Global Matching for Video Frame Interpolation with Large MotionChunxu Liu, Guozhen Zhang, Rui Zhao, Limin WangCVPR 2024 · 17 citations
- Event-based Motion Deblurring with Modality-Aware Decomposition and RecompositionWen Yang, Jinjian Wu, Leida Li, Weisheng Dong et al.ACM MM 2023 · 14 citations
