RT-BEV: Enhancing Real-Time BEV Perception for Autonomous Vehicles
Liangkai Liu, Jinkyu Lee, Kang G. Shin
Abstract
Vision-centric Bird’s Eye View (BEV) perception has become popular for enhancing the situational awareness of autonomous vehicles (AVs). It uses multiple cameras to create a 360° view, capturing essential details for the vehicle’s navigation and decision-making. However, reducing the end-to-end (e2e) BEV perception latency without sacrificing accuracy is challenging due to the lack of co-optimization of message communication and object detection. Prior work either compresses the dense detection model to reduce computation which can hurt accuracy and assume images are well synchronized, or focuses on worstcase communication delay without considering the characteristics of object detection. To meet this challenge, we propose RT-BEV, the first frame-work designed to co-optimize message communication and object detection to improve real-time e2e BEV perception without sacrificing accuracy. The main insight of RT-BEV lies in generating traffic environment- and context-aware Regions of Interest (ROIs) for AV safety, combined with ROI-aware message communication. RT-BEV features an ROI-aware Camera Synchronizer that adaptively determines message groups and allowable delays based on ROIs’ coverage. We also develop a ROIs Generator to model context-aware ROIs and a Feature Split & Merge component to handle variable-sized ROIs effectively. Furthermore, a Time Predictor forecasts timelines for processing ROIs, and a Coordinator jointly optimizes latency and accuracy for the entire e2e pipeline. We have implemented RT-BEV in a ROS-based BEV perception pipeline and evaluated it with the nuScenes dataset. RT-BEV is shown to significantly enhances real-time BEV perception, reducing average e2e latency by , maintaining high mean Average Precision (mAP), doubling the number of processed frames, and improving the frame efficiency score (FES) by compared to the existing approaches. Moreover, RT-BEV is shown to reduce the worst-case e2e latency by .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f2d86dc0-19c9-4c73-93bd-e05bee323745Builds on12
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- SparseBEV: High-Performance Sparse 3D Object Detection from Multi-Camera VideosHaisong Liu, Yao Teng, Tao Lu, Haiguang Wang et al.ICCV 2023 · 204 citations
- Prophet: Realizing a Predictable Real-time Perception Pipeline for Autonomous VehiclesLiangkai Liu, Zheng Dong, Yanzhi Wang, Weisong ShiRTSS 2022 · 34 citations
- QD-BEV : Quantization-aware View-guided Distillation for Multi-view 3D Object DetectionYifan Zhang, Zhen Dong, Huanrui Yang, Ming Lu et al.ICCV 2023 · 15 citations
- RT-MOT: Confidence-Aware Real-Time Scheduling Framework for Multi-Object Tracking TasksDonghwa Kang, Seunghoon Lee, Hoon Sung Chwa, Seung-Hwan Bae et al.RTSS 2022 · 10 citations
Related papers
- AsyncBEV: Cross-modal flow alignment in Asynchronous 3D Object DetectionShiming Wang, Holger Caesar, Liangliang Nan, Julian F. P. KooijICLR 2026 · 2 citations
- HotBEV: Hardware-oriented Transformer-based Multi-View 3D Detector for BEV PerceptionPeiyan Dong, Zhenglun Kong, Xin Meng, Pinrui Yu et al.NeurIPS 2023 · 6 citations
- BEVCooper: Accurate and Communication-Efficient Bird's-Eye-View Perception in Vehicular NetworksJiawei Hou, Peng Yang, Xiangxiang Dai, Mingliu Liu et al.INFOCOM 2026
- Asynchrony-Robust Collaborative Perception via Bird's Eye View FlowSizhe Wei, Yuxi Wei, Yue Hu, Yifan Lu et al.NeurIPS 2023 · 102 citations
- TBP-Former: Learning Temporal Bird's-Eye-View Pyramid for Joint Perception and Prediction in Vision-Centric Autonomous DrivingShaoheng Fang, Zi Wang, Yiqi Zhong, Junhao Ge et al.CVPR 2023
