DAIR-V2X: A Large-Scale Dataset for Vehicle-Infrastructure Cooperative 3D Object Detection
Haibao Yu, Yizhen Luo, Mao Shu, Yiyi Huo, Zebang Yang, Yifeng Shi, Zhenglong Guo, Hanyu Li, Xing Hu, Jirui Yuan, Zaiqing Nie
Abstract
Autonomous driving faces great safety challenges for a lack of global perspective and the limitation of long-range perception capabilities. It has been widely agreed that vehicle-infrastructure cooperation is required to achieve Level 5 autonomy. However, there is still NO dataset from real scenarios available for computer vision researchers to work on vehicle-infrastructure cooperation-related problems. To accelerate computer vision research and innovation for Vehicle-Infrastructure Cooperative Autonomous Driving (VICAD), we release DAIR-V2X Dataset, which is the first large-scale, multi-modality, multi-view dataset from real scenarios for VICAD. DAIR-V2X comprises 71254 LiDAR frames and 71254 Camera frames, and all frames are captured from real scenes with 3D annotations. The Vehicle-Infrastructure Cooperative 3D Object Detection problem (VIC3D) is introduced, formulating the problem of collaboratively locating and identifying 3D objects using sensory inputs from both vehicle and infrastructure. In addition to solving traditional 3D object detection problems, the solution of VIC3D needs to consider the temporal asynchrony problem between vehicle and infrastructure sensors and the data transmission cost between them. Furthermore, we propose Time Compensation Late Fusion (TCLF), a late fusion framework for the VIC3D task as a benchmark based on DAIR-V2X. Find data, code, and more up-to-date information at https://thudair.baai.ac.cn/index and https://github.com/AIR-Thu/dair-V2x.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 57ea2c8e-e96f-4e7a-923b-b1850ad9b7c7Cited by top-tier papers97
- Where2comm: Communication-Efficient Collaborative Perception via Spatial Confidence MapsYue Hu, Shaoheng Fang, Zixing Lei, Yiqi Zhong et al.NeurIPS 2022 · 537 citations
- How2comm: Communication-Efficient and Collaboration-Pragmatic Multi-Agent PerceptionDingkang Yang, Kun Yang, Yuzheng Wang, Jing Liu et al.NeurIPS 2023 · 160 citations
- DeepAccident: A Motion and Accident Prediction Benchmark for V2X Autonomous DrivingTianqi Wang, Sukmin Kim, Wenxuan Ji, Enze Xie et al.AAAI 2024 · 132 citations
- Rope3D: The Roadside Perception Dataset for Autonomous Driving and Monocular 3D Object Detection TaskXiaoqing Ye, Mao Shu, Hanyu Li, Yifeng Shi et al.CVPR 2022 · 130 citations
- An Extensible Framework for Open Heterogeneous Collaborative PerceptionYifan Lu, Yue Hu, Yiqi Zhong, Dequan Wang et al.ICLR 2024 · 116 citations
Builds on7
- Learning Distilled Collaboration Graph for Multi-Agent PerceptionYiming Li, Shunli Ren, Pengxiang Wu, Siheng Chen et al.NeurIPS 2021 · 464 citations
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora et al.CVPR 2020
- PointPainting: Sequential Fusion for 3D Object DetectionSourabh Vora, Alex H. Lang, Bassam Helou, Oscar BeijbomCVPR 2020
- Scalability in Perception for Autonomous Driving: Waymo Open DatasetPei Sun, Henrik Kretzschmar, Xerxes Dotiwalla, Aurelien Chouard et al.CVPR 2020
- Categorical Depth Distribution Network for Monocular 3D Object DetectionCody Reading, Ali Harakeh, Julia Chae, Steven L. WaslanderCVPR 2021
Related papers
- V2V4Real: A Real-World Large-Scale Dataset for Vehicle-to-Vehicle Cooperative PerceptionRunsheng Xu, Xin Xia, Jinlong Li, Hanzhao Li et al.CVPR 2023
- Flow-Based Feature Fusion for Vehicle-Infrastructure Cooperative 3D Object DetectionHaibao Yu, Yingjuan Tang, Enze Xie, Jilei Mao et al.NeurIPS 2023 · 76 citations
- V2U4Real: A Real-world Large-scale Dataset for Vehicle-to-UAV Cooperative PerceptionWeijia Li, Haoen Xiang, Tianxu Wang, Shuaibing Wu et al.CVPR 2026 · 4 citations
- HoloVic: Large-scale Dataset and Benchmark for Multi-Sensor Holographic Intersection and Vehicle-Infrastructure CooperativeCong Ma, Lei Qiao, Chengkai Zhu, Kai Liu et al.CVPR 2024 · 14 citations
- V2X-Seq: A Large-Scale Sequential Dataset for Vehicle-Infrastructure Cooperative Perception and ForecastingHaibao Yu, Wenxian Yang, Hongzhi Ruan, Zhenwei Yang et al.CVPR 2023
