Retinomorphic Object Detection in Asynchronous Visual Streams
Jianing Li, Xiao Wang, Lin Zhu, Jia Li, Tiejun Huang, Yonghong Tian
摘要
Due to high-speed motion blur and challenging illumination, conventional frame-based cameras have encountered an important challenge in object detection tasks. Neuromorphic cameras that output asynchronous visual streams instead of intensity frames, by taking the advantage of high temporal resolution and high dynamic range, have brought a new perspective to address the challenge. In this paper, we propose a novel problem setting, retinomorphic object detection, which is the first trial that integrates foveal-like and peripheral-like visual streams. Technically, we first build a large-scale multimodal neuromorphic object detection dataset (i.e., PKU-Vidar-DVS) over 215.5k spatio-temporal synchronized labels. Then, we design temporal aggregation representations to preserve the spatio-temporal information from asynchronous visual streams. Finally, we present a novel bio-inspired unifying framework to fuse two sensing modalities via a dynamic interaction mechanism. Our experimental evaluation shows that our approach has significant improvements over the state-of-the-art methods with the single-modality, especially in high-speed motion and low-light scenarios. We hope that our work will attract further research into this newly identified, yet crucial research direction. Our dataset can be available at https://www.pkuml.org/resources/pku-vidar-dvs.html.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Deep Directly-Trained Spiking Neural Networks for Object DetectionQiaoyi Su, Yuhong Chou, Yifan Hu, Jianing Li 等ICCV 2023 · 被引用 143 次
- HARDVS: Revisiting Human Activity Recognition with Dynamic Vision SensorsXiao Wang, Zongzhen Wu, Bo Jiang, Zhimin Bao 等AAAI 2024 · 被引用 80 次
- Learning Graph-embedded Key-event Back-tracing for Object Tracking in Event CloudsZhiyu Zhu, Junhui Hou, Xianqiang LyuNeurIPS 2022 · 被引用 48 次
- Optical Flow for Spike Camera with Hierarchical Spatial-Temporal Spike FusionRui Zhao, Ruiqin Xiong, Jian Zhang, Xinfeng Zhang 等AAAI 2024 · 被引用 23 次
- DMR: Decomposed Multi-Modality Representations for Frames and Events Fusion in Visual Reinforcement LearningHaoran Xu, Peixi Peng, Guang Tan, Yuan Li 等CVPR 2024 · 被引用 5 次
它引用的顶会 Paper8
- Going Deeper With Directly-Trained Larger Spiking Neural NetworksHanle Zheng, Yujie Wu, Lei Deng, Yifan Hu 等AAAI 2021 · 被引用 694 次
- End-to-End Learning of Representations for Asynchronous Event-Based DataDaniel Gehrig, Antonio Loquercio, Konstantinos G. Derpanis, Davide ScaramuzzaICCV 2019 · 被引用 427 次
- Learning to Detect Objects with a 1 Megapixel Event CameraEtienne Perot, Pierre de Tournemire, Davide Nitti, Jonathan Masci 等NeurIPS 2020 · 被引用 381 次
- Deep Multimodal Fusion by Channel ExchangingYikai Wang, Wenbing Huang, Fuchun Sun, Tingyang Xu 等NeurIPS 2020 · 被引用 321 次
- NeuSpike-Net: High Speed Video Reconstruction via Bio-inspired Neuromorphic CamerasLin Zhu, Jianing Li, Xiao Wang, Tiejun Huang 等ICCV 2021 · 被引用 55 次
相关 Paper
- Recognizing High-Speed Moving Objects with Spike CameraJunwei Zhao, Jianming Ye, Shiliang Zhang, Zhaofei Yu 等ACM MM 2023 · 被引用 4 次
- 1000 FPS HDR Video with a Spike-RGB Hybrid CameraYakun Chang, Chu Zhou, Yuchen Hong, Liwen Hu 等CVPR 2023
- Intensity-Robust Autofocus for Spike CameraChangqing Su, Zhiyuan Ye, Yongsheng Xiao, You Zhou 等CVPR 2024 · 被引用 2 次
- "Seeing" Electric Network Frequency from EventsLexuan Xu, Guang Hua, Haijian Zhang, Lei Yu 等CVPR 2023
- Recognizing Ultra-High-Speed Moving Objects with Bio-Inspired Spike CameraJunwei Zhao, Shiliang Zhang, Zhaofei Yu, Tiejun HuangAAAI 2024 · 被引用 5 次
