Spk2VidNet: A Hierarchical Recurrent Architecture for High-Fidelity Video Reconstruction from Long Spike-Camera Streams
Yuanlin Wang, Ruiqin Xiong, Jiyu Xie, Zhenkun Zhu, Zhaofei Yu, Xiaopeng Fan, Tiejun Huang
Abstract
Spike camera is a neuromorphic vision sensor with ultra-high temporal resolution, capable of capturing fast-moving scenes by firing a stream of binary spikes. However, its relatively low spatial resolution limits the acquisition of fine-grained visual details, motivating research on spike camera super resolution (SCSR). Existing SCSR methods typically operate on fixed-length spike sequences, where the accessible information is confined to a local temporal neighborhood. Moreover, spike fluctuations hinder intensity information extraction. Both factors affect the performance of SCSR. To address these issues, we propose a hierarchical recurrent network named Spk2VidNet to reconstruct high-fidelity high resolution image sequences from low resolution spike data. To mitigate fluctuations, Spk2VidNet progressively exploits temporal correlations within spike stream to enhance feature representation by hierarchically enlarging temporal receptive fields. Within recurrent phase, we introduce an alignment module that leverages the motion consistency among multiple frames to jointly estimate and mutually refine inter-frame motions, achieving more accurate temporal alignment. In addition, we propose a fusion module to adaptively integrate neighboring aligned features based on multi-scale similarity for robust feature aggregation. We further propose a segment-wise training with state transfer strategy to efficiently model long-term dependencies with limited GPU memory, thereby leveraging rich subpixel cues for improved super resolution. Experiments on synthetic and real-captured spike data demonstrate that Spk2VidNet achieves state-of-the-art performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dc933ba6-3d70-4f96-9ff7-39cf3aff84f0Builds on14
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 522 citations
- Rethinking Alignment in Video Super-Resolution TransformersShuwei Shi, Jinjin Gu, Liangbin Xie, Xintao Wang et al.NeurIPS 2022 · 134 citations
- Super Resolve Dynamic Scene from Continuous Spike StreamsJing Zhao, Jiyu Xie, Ruiqin Xiong, Jian Zhang et al.ICCV 2021 · 42 citations
- Learning Temporal-Ordered Representation for Spike Streams Based on Discrete Wavelet TransformsJiyuan Zhang, Shanshan Jia, Zhaofei Yu, Tiejun HuangAAAI 2023 · 33 citations
- Learning to Super-resolve Dynamic Scenes for Neuromorphic Spike CameraJing Zhao, Ruiqin Xiong, Jian Zhang, Rui Zhao et al.AAAI 2023 · 21 citations
Related papers
- Spk2SRImgNet: Super-Resolve Dynamic Scene from Spike Stream via Motion Aligned Collaborative FilteringYuanlin Wang, Yiyang Zhang, Ruiqin Xiong, Jing Zhao et al.CVPR 2025
- Spk2ImgNet: Learning To Reconstruct Dynamic Scene From Continuous Spike StreamJing Zhao, Ruiqin Xiong, Hangfan Liu, Jian Zhang et al.CVPR 2021
- Spatio-Temporal Interactive Learning for Efficient Image Reconstruction of Spiking CamerasBin Fan, Jiaoyang Yin, Yuchao Dai, Chao Xu et al.NeurIPS 2024 · 7 citations
- Super-Resolution Reconstruction from Bayer-Pattern Spike StreamsYanchen Dong, Ruiqin Xiong, Jian Zhang, Zhaofei Yu et al.CVPR 2024
- Optical Flow for Spike Camera with Hierarchical Spatial-Temporal Spike FusionRui Zhao, Ruiqin Xiong, Jian Zhang, Xinfeng Zhang et al.AAAI 2024 · 23 citations
