V2V: Scaling Event-Based Vision through Efficient Video-to-Voxel Simulation
Hanyue Lou, Jinxiu Liang, Minggui Teng, Yi Wang, Boxin Shi
摘要
Event-based cameras offer unique advantages such as high temporal resolution, high dynamic range, and low power consumption. However, the massive storage requirements and I/O burdens of existing synthetic data generation pipelines and the scarcity of real data prevent event-based training datasets from scaling up, limiting the development and generalization capabilities of event vision models. To address this challenge, we introduce Video-to-Voxel (V2V), an approach that directly converts conventional video frames into event-based voxel grid representations, bypassing the storage-intensive event stream generation entirely. V2V enables a 150 times reduction in storage requirements while supporting on-the-fly parameter randomization for enhanced model robustness. Leveraging this efficiency, we train several video reconstruction and optical flow estimation model architectures on 10,000 diverse videos totaling 52 hours--an order of magnitude larger than existing event datasets, yielding substantial improvements.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- AE2VID: Event-based Video Reconstruction via Aperture ModulationChenxu Bai, Boyu Li, Peiqi Duan, Xinyu Zhou 等CVPR 2026 · 被引用 1 次
- Texvent: Asynchronous Event Data Simulation via Text PromptRuofei Wang, Peiqi Duan, Ka Chun Cheung, Simon See 等CVPR 2026
它引用的顶会 Paper10
- Frozen in Time: A Joint Video and Image Encoder for End-to-End RetrievalMax Bain, Arsha Nagrani, Gül Varol, Andrew ZissermanICCV 2021 · 被引用 1,550 次
- Event-based Video Reconstruction Using TransformerWenming Weng, Yueyi Zhang, Zhiwei XiongICCV 2021 · 被引用 139 次
- Event-based Video Reconstruction via Potential-assisted Spiking Neural NetworkLin Zhu, Xiao Wang, Yi Chang, Jianing Li 等CVPR 2022 · 被引用 109 次
- Spatio-Temporal Recurrent Networks for Event-Based Optical Flow EstimationZiluo Ding, Rui Zhao, Jiyuan Zhang, Tianxiao Gao 等AAAI 2022 · 被引用 76 次
- Learning Optical Flow from Event Camera with Rendered DatasetXinglong Luo, Kunming Luo, Ao Luo, Zhengning Wang 等ICCV 2023 · 被引用 28 次
相关 Paper
- Video to Events: Recycling Video Datasets for Event CamerasDaniel Gehrig, Mathias Gehrig, Javier Hidalgo-Carrió, Davide ScaramuzzaCVPR 2020
- How to Learn a Domain-Adaptive Event Simulator?Daxin Gu, Jia Li, Yu Zhang, Yonghong TianACM MM 2021 · 被引用 8 次
- End-to-End Learning of Representations for Asynchronous Event-Based DataDaniel Gehrig, Antonio Loquercio, Konstantinos G. Derpanis, Davide ScaramuzzaICCV 2019 · 被引用 427 次
- E2PNet: Event to Point Cloud Registration with Spatio-Temporal Representation LearningXiuhong Lin, Changjie Qiu, Zhipeng Cai, Siqi Shen 等NeurIPS 2023 · 被引用 18 次
- A Voxel Graph CNN for Object Classification with Event CamerasYongjian Deng, Hao Chen, Hai Liu, Youfu LiCVPR 2022 · 被引用 55 次
