Toward Real-world BEV Perception: Depth Uncertainty Estimation via Gaussian Splatting
Shu-Wei Lu, Yi-Hsuan Tsai, Yi-Ting Chen
Abstract
Bird's-eye view (BEV) perception has gained significant attention because it provides a unified representation to fuse multiple view images and enables a wide range of downstream autonomous driving tasks, such as forecasting and planning. Recent state-of-the-art models utilize projectionbased methods which formulate BEV perception as query learning to bypass explicit depth estimation. While we observe promising advancements in this paradigm, they still fall short of real-world applications because of the lack of uncertainty modeling and expensive computational requirement. In this work, we introduce GaussianLSS, an uncertainty-aware BEV perception framework that revisits the unprojection-based method, specifically the Lift-Splat-Shoot (LSS) paradigm, and enhances it with depth uncertainty modeling. Our GaussianLSS represents spatial dispersion by learning a soft depth mean and computing the variance of the depth distribution, which implicitly captures object extents. We then transform the depth distribution into 3D Gaussians and rasterize them to construct uncertainty-aware BEV features. We evaluate GaussianLSS on the nuScenes dataset, achieving state-of-the-art performance compared to unprojection-based methods. In particular, it provides significant advantages in speed, running 2x faster, and in memory efficiency, using 0.3x less memory compared to projection-based methods, while achieving competitive performance with only a 0.7% IoU difference. See our project page for more details: https://hcis-lab.github.io/GaussianLSS/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- RCTDistill: Cross-Modal Knowledge Distillation Framework for Radar-Camera 3D Object Detection with Temporal FusionGeonho Bang, Minjae Seong, Jisong Kim, Geunju Baek et al.ICCV 2025 · 6 citations
- DLWM: Dual Latent World Models enable Holistic Gaussian-centric Pre-training in Autonomous DrivingYiyao Zhu, Ying Xue, Haiming Zhang, Guangfeng Jiang et al.CVPR 2026 · 2 citations
- Benchmarking PhD-Level Coding in 3D Geometric Computer VisionWenyi Li, Renkai Luo, Yue Yu, Huan-ang Gao et al.CVPR 2026 · 2 citations
- GSV2X: Geometry-Aware Uncertainty Modeling and Orthogonal Fusion for Robust Roadside PerceptionJianqiang Xu, Gensheng Pei, Huafeng Liu, Yazhou YaoCVPR 2026 · 2 citations
- Towards 3D Object-Centric Feature Learning for Semantic Scene CompletionWeihua Wang, Yubo Cui, Xiangru Lin, Zhiheng Li et al.AAAI 2026
Builds on19
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- BEVDepth: Acquisition of Reliable Depth for Multi-View 3D Object DetectionYinhao Li, Zheng Ge, Guanyi Yu, Jinrong Yang et al.AAAI 2023 · 954 citations
- PETRv2: A Unified Framework for 3D Perception from Multi-Camera ImagesYingfei Liu, Junjie Yan, Fan Jia, Shuailin Li et al.ICCV 2023 · 513 citations
- Cross-view Transformers for real-time Map-view Semantic SegmentationBrady Zhou, Philipp KrähenbühlCVPR 2022 · 279 citations
- FB-BEV: BEV Representation from Forward-Backward View TransformationsZhiqi Li, Zhiding Yu, Wenhai Wang, Anima Anandkumar et al.ICCV 2023 · 144 citations
Related papers
- GaussianFusion: Unified 3D Gaussian Representation for Multi-Modal Fusion PerceptionXiao Zhao, Chang Liu, Mingxu Zhu, Zheyuan Zhang et al.ICLR 2026 · 1 citation
- Parametric Depth Based Feature Representation Learning for Object Detection and Segmentation in Bird's-Eye ViewJiayu Yang, Enze Xie, Miaomiao Liu, José M. ÁlvarezICCV 2023 · 9 citations
- SQS: Enhancing Sparse Perception Models via Query-based Splatting in Autonomous DrivingHaiming Zhang, Yiyao Zhu, Wending Zhou, Xu Yan et al.NeurIPS 2025 · 5 citations
- Instance-Aware Multi-Camera 3D Object Detection with Structural Priors Mining and Self-Boosting LearningYang Jiao, Zequn Jie, Shaoxiang Chen, Lechao Cheng et al.AAAI 2024 · 13 citations
- OccluBEV: Occlusion Aware Spatiotemporal Modeling for Multi-view 3D Object DetectionZiteng Wen, Hai Xu, Chenyu Liu, Tao Guo et al.ACM MM 2023 · 5 citations
