SparseLaneSTP: Leveraging Spatio-Temporal Priors with Sparse Transformers for 3D Lane Detection
Maximilian Pittner, Joel Janai, Mario Faigle, Alexandru Paul Condurache
Abstract
3D lane detection has emerged as a critical challenge in autonomous driving, encompassing identification and localization of lane markings and the 3D road surface. Conventional 3D methods detect lanes from dense birds-eye-viewed (BEV) features, though erroneous transformations often result in a poor feature representation misaligned with the true 3D road surface. While recent sparse lane detectors have surpassed dense BEV approaches, they completely disregard valuable lane-specific priors. Furthermore, existing methods fail to utilize historic lane observations, which yield the potential to resolve ambiguities in situations of poor visibility. To address these challenges, we present SparseLaneSTP, a novel method that integrates both geometric properties of the lane structure and temporal information into a sparse lane transformer. It introduces a new lane-specific spatio-temporal attention mechanism, a continuous lane representation tailored for sparse architectures as well as temporal regularization. Identifying weaknesses of existing 3D lane datasets, we also introduce a precise and consistent 3D lane dataset using a simple yet effective auto-labeling strategy. Our experimental section proves the benefits of our contributions and demonstrates state-of-the-art performance across all detection and error metrics on existing 3D lane detection benchmarks as well as on our novel dataset.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 04ede73d-62b9-4f29-910e-6e14d80762b2Builds on22
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra et al.NeurIPS 2022 · 5,493 citations
- FlashAttention-2: Faster Attention with Better Parallelism and Work PartitioningTri DaoICLR 2024 · 2,600 citations
- Learning Lightweight Lane Detection CNNs by Self Attention DistillationYuenan Hou, Zheng Ma, Chunxiao Liu, Chen Change LoyICCV 2019 · 666 citations
- An End-to-End Transformer Model for 3D Object DetectionIshan Misra, Rohit Girdhar, Armand JoulinICCV 2021 · 602 citations
Related papers
- LaneCPP: Continuous 3D Lane Detection Using Physical PriorsMaximilian Pittner, Joel Janai, Alexandru Paul ConduracheCVPR 2024 · 24 citations
- PVALane: Prior-Guided 3D Lane Detection with View-Agnostic Feature AlignmentZewen Zheng, Xuemin Zhang, Yongqiang Mou, Xiang Gao et al.AAAI 2024 · 26 citations
- BEV-LaneDet: An Efficient 3D Lane Detection Based on Virtual Camera via Key-PointsRuihao Wang, Jian Qin, Kaiying Li, Yaochen Li et al.CVPR 2023
- LATR: 3D Lane Detection from Monocular Images with TransformerYueru Luo, Chaoda Zheng, Xu Yan, Tang Kun et al.ICCV 2023 · 69 citations
- Rethinking Lanes and Points in Complex Scenarios for Monocular 3D Lane DetectionYifan Chang, Junjie Huang, Xiaofeng Wang, Yun Ye et al.CVPR 2025
