DV-3DLane: End-to-end Multi-modal 3D Lane Detection with Dual-view Representation
Yueru Luo, Shuguang Cui, Zhen Li
Abstract
Accurate 3D lane estimation is crucial for ensuring safety in autonomous driving. However, prevailing monocular techniques suffer from depth loss and lighting variations, hampering accurate 3D lane detection. In contrast, LiDAR points offer geometric cues and enable precise localization. In this paper, we present DV-3DLane, a novel end-to-end Dual-View multi-modal 3D Lane detection framework that synergizes the strengths of both images and LiDAR points. We propose to learn multi-modal features in dual-view spaces, i.e., perspective view (PV) and bird's-eye-view (BEV), effectively leveraging the modal-specific information. To achieve this, we introduce three designs: 1) A bidirectional feature fusion strategy that integrates multi-modal features into each view space, exploiting their unique strengths. 2) A unified query generation approach that leverages lane-aware knowledge from both PV and BEV spaces to generate queries. 3) A 3D dual-view deformable attention mechanism, which aggregates discriminative features from both PV and BEV spaces into queries for accurate 3D lane detection. Extensive experiments on the public benchmark, OpenLane, demonstrate the efficacy and efficiency of DV-3DLane. It achieves state-of-the-art performance, with a remarkable 11.2 gain in F1 score and a substantial 53.5% reduction in errors. The code is available at https://github.com/JMoonr/dv-3dlane.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 271fc061-7e4e-4731-9bd9-812077072fa1Cited by top-tier papers2
- ReManNet: A Riemannian Manifold Network for Monocular 3D Lane DetectionChengzhi Hong, Bijun LiCVPR 2026
- GLane3D: Detecting Lanes with Graph of 3D KeypointsHalil Ibrahim Öztürk, Muhammet Esat Kalfaoglu, Ozsel KilincCVPR 2025
Builds on27
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- BEVDepth: Acquisition of Reliable Depth for Multi-View 3D Object DetectionYinhao Li, Zheng Ge, Guanyi Yu, Jinrong Yang et al.AAAI 2023 · 954 citations
- TransFusion: Robust LiDAR-Camera Fusion for 3D Object Detection with TransformersXuyang Bai, Zeyu Hu, Xinge Zhu, Qingqiu Huang et al.CVPR 2022 · 794 citations
- BEVFusion: A Simple and Robust LiDAR-Camera Fusion FrameworkTingting Liang, Hongwei Xie, Kaicheng Yu, Zhongyu Xia et al.NeurIPS 2022 · 762 citations
- Learning Lightweight Lane Detection CNNs by Self Attention DistillationYuenan Hou, Zheng Ma, Chunxiao Liu, Chen Change LoyICCV 2019 · 666 citations
Related papers
- PVALane: Prior-Guided 3D Lane Detection with View-Agnostic Feature AlignmentZewen Zheng, Xuemin Zhang, Yongqiang Mou, Xiang Gao et al.AAAI 2024 · 26 citations
- LaneCMKT: Boosting Monocular 3D Lane Detection with Cross-Modal Knowledge TransferRunkai Zhao, Heng Wang, Weidong CaiACM MM 2024 · 4 citations
- LATR: 3D Lane Detection from Monocular Images with TransformerYueru Luo, Chaoda Zheng, Xu Yan, Tang Kun et al.ICCV 2023 · 69 citations
- Anchor3DLane: Learning to Regress 3D Anchors for Monocular 3D Lane DetectionShaofei Huang, Zhenwei Shen, Zehao Huang, Zi-han Ding et al.CVPR 2023
- VISTA: Boosting 3D Object Detection via Dual Cross-VIew SpaTial AttentionShengheng Deng, Zhihao Liang, Lin Sun, Kui JiaCVPR 2022 · 92 citations
