240FPS Stereo Vision from Monocular Mixed Spikes
Yeliduosi Xiaokaiti, Yakun Chang, Yang Bai, Zhaojun Huang, Peiqi Duan, Boxin Shi
摘要
Stereo vision is fundamental for enabling machines to perceive and interact with the world. While monocular stereo methods offer hardware compactness, they struggle with generalization due to reliance on data-driven priors. Binocular and multi-view systems improve accuracy but incur higher hardware complexity and data inefficiency. In this paper, we introduce a monocular solution for high-framerate stereo vision via temporal optical modulation. The modulation directs light from two views onto a single sensor in a mixed manner, while periodically attenuating one view at 60 Hz. To capture the temporal variations introduced by this modulation, we employ a high-speed spike camera that records the mixed scene as temporally dense spikes. The high temporal resolution of these spikes enables the construction of a linear system for efficient binocular video decoupling. Consequently, we introduce a two-stage decoding methodology for achieving high-quality stereo vision: An efficient least-squares-based baseline reconstruction followed by a deep learning refinement module. Experimental results demonstrate that our approach achieves 240FPS binocular video reconstruction with superior accuracy compared to monocular systems, while maintaining the hardware compactness and data efficiency. Code is available at https://github.com/yongqiye00/MonoSpikeStereo.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper12
- Depth Anything V2Lihe Yang, Bingyi Kang, Zilong Huang, Zhen Zhao 等NeurIPS 2024 · 被引用 2,305 次
- Practical Stereo Matching via Cascaded Recurrent Network with Adaptive CorrelationJiankun Li, Peisen Wang, Pengfei Xiong, Tao Cai 等CVPR 2022 · 被引用 294 次
- Two-in-One Depth: Bridging the Gap Between Monocular and Binocular Self-supervised Depth EstimationZhengming Zhou, Qiulei DongICCV 2023 · 被引用 16 次
- Depth Pro: Sharp Monocular Metric Depth in Less Than a SecondAlexey Bochkovskiy, Amaël Delaunoy, Hugo Germain, Marcel Santos 等ICLR 2025 · 被引用 15 次
- Spatio-Temporal Interactive Learning for Efficient Image Reconstruction of Spiking CamerasBin Fan, Jiaoyang Yin, Yuchao Dai, Chao Xu 等NeurIPS 2024 · 被引用 7 次
相关 Paper
- Enhancing Motion Deblurring in High-Speed Scenes with Spike StreamsShiyan Chen, Jiyuan Zhang, Yajing Zheng, Tiejun Huang 等NeurIPS 2023 · 被引用 21 次
- BulletTime4D: Towards High Spatio-Temporal Resolution Dynamic Scene Rendering via Spike-Guided Stereo VisionYiqian Chang, Haoran Xu, Qinghong Ye, Jianing Li 等AAAI 2026
- SpikeStereoNet: A Brain-Inspired Framework for Stereo Depth Estimation from Spike StreamsZhuoheng Gao, Yihao Li, Jiyao Zhang, Rui Zhao 等ICLR 2026 · 被引用 2 次
- Optical Flow Estimation for Spiking CameraLiwen Hu, Rui Zhao, Ziluo Ding, Lei Ma 等CVPR 2022 · 被引用 48 次
- HFR and HDR Video from Multi-Attenuated Spikes Using a Rapidly Rotating SpokeND FilterYakun Chang, Zhaojun Huang, Siqi Yang, Yeliduosi Xiaokaiti 等CVPR 2026
