Asynchronous Collaborative Graph Representation for Frames and Events
Dianze Li, Jianing Li, Xu Liu, Xiaopeng Fan, Yonghong Tian
摘要
Integrating frames and events has become a widely accepted solution for various tasks in challenging scenarios. However, most multimodal methods directly convert events into image-like formats synchronized with frames and process each stream through separate two-branch backbones, making it difficult to fully exploit the spatiotemporal events while limiting inference frequency to the frame rate. To address these problems, we propose a novel asynchronous collaborative graph representation, namely ACGR, which is the first trial to explore a unified graph framework for asynchronously processing frames and events with high performance and low latency. Technically, we first construct unimodal graphs for frames and events to preserve their spatiotemporal properties and sparsity. Then, an asynchronous collaborative alignment module is designed to align and fuse frames and events into a unified graph and the ACGR is generated through graph convolutional networks. Finally, we innovatively introduce domain adaptation to enable cross-modal interactions between frames and events by aligning their feature spaces. Experimental results show that our approach outperforms state-of-the-art methods in both object detection and depth estimation tasks, while significantly reducing computational latency and achieving real-time inference up to 200 Hz. Our code can be available at https://github.com/dianzl/ACGR .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Maximizing Asynchronicity in Event-based Neural NetworksHaiqing Hao, Nikola Zubic, Weihua He, Zhipeng Sui 等ICLR 2026 · 被引用 2 次
- Beyond Duality: A Hybrid Framework of Leveraging Shared and Private Features for RGB-Event Object DetectionKeyao Wang, Shuai Liu, Hengda Shi, Lukui Shi 等CVPR 2026
它引用的顶会 Paper20
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- End-to-End Learning of Representations for Asynchronous Event-Based DataDaniel Gehrig, Antonio Loquercio, Konstantinos G. Derpanis, Davide ScaramuzzaICCV 2019 · 被引用 427 次
- Deep Directly-Trained Spiking Neural Networks for Object DetectionQiaoyi Su, Yuhong Chou, Yifan Hu, Jianing Li 等ICCV 2023 · 被引用 143 次
- Object Tracking by Jointly Exploiting Frame and Event DomainJiqing Zhang, Xin Yang, Yingkai Fu, Xiaopeng Wei 等ICCV 2021 · 被引用 141 次
相关 Paper
- AEGNN: Asynchronous Event-based Graph Neural NetworksSimon Schaefer, Daniel Gehrig, Davide ScaramuzzaCVPR 2022 · 被引用 135 次
- Ev-3DOD: Pushing the Temporal Boundaries of 3D Object Detection with Event CamerasHoonhee Cho, Jae-Young Kang, Youngho Kim, Kuk-Jin YoonCVPR 2025
- Graph Neural Network Combining Event Stream and Periodic Aggregation for Low-Latency Event-based VisionManon Dampfhoffer, Thomas Mesquida, Damien Joubert, Thomas Dalgaty 等CVPR 2025
- AIMDepth: Asymmetric Image-Event Mamba for Monocular Depth EstimationLuoxi Jing, Dianxi Shi, YuShe Cao, Yuanze Wang 等CVPR 2026
- When Every Millisecond Counts: Real-Time Anomaly Detection via the Multimodal Asynchronous Hybrid NetworkDong Xiao, Guangyao Chen, Peixi Peng, Yangru Huang 等ICML 2025
