VIPS: real-time perception fusion for infrastructure-assisted autonomous driving
Shuyao Shi, Jiahe Cui, Zhehao Jiang, Zhenyu Yan, Guoliang Xing, Jianwei Niu, Zhenchao Ouyang
Abstract
Infrastructure-assisted autonomous driving is an emerging paradigm that expects to significantly improve the driving safety of autonomous vehicles. The key enabling technology for this vision is to fuse LiDAR results from the roadside infrastructure and the vehicle to improve the vehicle's perception in real time. In this work, we propose VIPS, a novel lightweight system that can achieve decimeter-level and real-time (up to 100 ms) perception fusion between driving vehicles and roadside infrastructure. The key idea of VIPS is to exploit highly efficient matching of graph structures that encode objects' lean representations as well as their relationships, such as locations, semantics, sizes, and spatial distribution. Moreover, by leveraging the tracked motion trajectories, VIPS can maintain the spatial and temporal consistency of the scene, which effectively mitigates the impact of asynchronous data frames and unpredictable communication/compute delays. We implement VIPS end-to-end based on a campus smart lamppost testbed. To evaluate the performance of VIPS under diverse situations, we also collect two new multi-view point cloud datasets using the smart lamppost testbed and an autonomous driving simulator, respectively. Experiment results show that VIPS can extend the vehicle's perception range by 140% within 58 ms on average, and delivers a 4X improvement in perception fusion accuracy and 47X data transmission saving over existing approaches. A video demo of VIPS based on the lamppost dataset is available at https://youtu.be/zW4oi_EWOu0.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 19826ab6-efe7-4c1c-9b4f-b252dfcdaee1Cited by top-tier papers26
- A Workload-Aware DVFS Robust to Concurrent Tasks for Mobile DevicesChengdong Lin, Kun Wang, Zhenjiang Li, Yu PuMobiCom 2023 · 52 citations
- Robust Real-time Multi-vehicle Collaboration on Asynchronous SensorsQingzhao Zhang, Xumiao Zhang, Ruiyang Zhu, Fan Bai et al.MobiCom 2023 · 44 citations
- VI-Map: Infrastructure-Assisted Real-Time HD Mapping for Autonomous DrivingYuze He, Chen Bian, Jingfei Xia, Shuyao Shi et al.MobiCom 2023 · 35 citations
- On Data Fabrication in Collaborative Vehicular Perception: Attacks and CountermeasuresQingzhao Zhang, Shuowei Jin, Ruiyang Zhu, Jiachen Sun et al.USENIX Security 2024 · 25 citations
- Soar: Design and Deployment of A Smart Roadside Infrastructure System for Autonomous DrivingShuyao Shi, Neiwen Ling, Zhehao Jiang, Xuan Huang et al.MobiCom 2024 · 25 citations
Builds on7
- STD: Sparse-to-Dense 3D Object Detector for Point CloudZetong Yang, Yanan Sun, Shu Liu, Xiaoyong Shen et al.ICCV 2019 · 840 citations
- EMP: edge-assisted multi-vehicle perceptionXumiao Zhang, Anlan Zhang, Jiachen Sun, Xiao Zhu et al.MobiCom 2021 · 137 citations
- VI-eye: semantic-based 3D point cloud registration for infrastructure-assisted autonomous drivingYuze He, Li Ma, Zhehao Jiang, Yi Tang et al.MobiCom 2021 · 76 citations
- Demystifying millimeter-wave V2X: towards robust and efficient directional connectivity under high mobilitySong Wang, Jingqi Huang, Xinyu ZhangMobiCom 2020 · 47 citations
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora et al.CVPR 2020
Related papers
- VI-Planning: Infrastructure-Assisted Real-Time Planning Optimization for Autonomous DrivingYang Lu, Jie Wang, Xiaoyun Dong, Ziyao Huang et al.MobiCom 2025 · 1 citation
- VILAM: Infrastructure-assisted 3D Visual Localization and Mapping for Autonomous DrivingJiahe Cui, Shuyao Shi, Yuze He, Jianwei Niu et al.NSDI 2024 · 17 citations
- Improving Multi-Vehicle Perception Fusion with Millimeter-Wave Radar AssistanceZhiqing Luo, Yi Wang, Yingying He, Wei WangINFOCOM 2025 · 8 citations
- FARFusion V2: A Geometry-based Radar-Camera Fusion Method on the Ground for Roadside Far-Range 3D Object DetectionYao Li, Jiajun Deng, Yuxuan Xiao, Yingjie Wang et al.ACM MM 2024 · 5 citations
- DAIR-V2X: A Large-Scale Dataset for Vehicle-Infrastructure Cooperative 3D Object DetectionHaibao Yu, Yizhen Luo, Mao Shu, Yiyi Huo et al.CVPR 2022 · 475 citations
