Laneformer: Object-Aware Row-Column Transformers for Lane Detection
Jianhua Han, Xiajun Deng, Xinyue Cai, Zhen Yang, Hang Xu, Chunjing Xu, Xiaodan Liang
Abstract
We present Laneformer, a conceptually simple yet powerful transformer-based architecture tailored for lane detection that is a long-standing research topic for visual perception in autonomous driving. The dominant paradigms rely on purely CNN-based architectures which often fail in incorporating relations of long-range lane points and global contexts induced by surrounding objects (e.g., pedestrians, vehicles). Inspired by recent advances of the transformer encoder-decoder architecture in various vision tasks, we move forwards to design a new end-to-end Laneformer architecture that revolutionizes the conventional transformers into better capturing the shape and semantic characteristics of lanes, with minimal overhead in latency. First, coupling with deformable pixel-wise self-attention in the encoder, Laneformer presents two new row and column self-attention operations to efficiently mine point context along with the lane shapes. Second, motivated by the appearing objects would affect the decision of predicting lane segments, Laneformer further includes the detected object instances as extra inputs of multi-head attention blocks in the encoder and decoder to facilitate the lane point detection by sensing semantic contexts. Specifically, the bounding box locations of objects are added into Key module to provide interaction with each pixel and query while the ROI-aligned features are inserted into Value module. Extensive experiments demonstrate our Laneformer achieves state-of-the-art performances on CULane benchmark, in terms of 77.1% F1 score. We hope our simple and effective Laneformer will serve as a strong baseline for future research in self-attention models for lane detection.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9cb58bec-b856-493c-bef7-564b557080dfCited by top-tier papers7
- ADNet: Lane Shape Prediction via Anchor DecompositionLingyu Xiao, Xiang Li, Sen Yang, Wankou YangICCV 2023 · 56 citations
- Lane2Seq: Towards Unified Lane Detection via Sequence GenerationKunyang ZhouCVPR 2024 · 28 citations
- Sketch and Refine: Towards Fast and Accurate Lane DetectionChao Chen, Jie Liu, Chang Zhou, Jie Tang et al.AAAI 2024 · 27 citations
- Visual Traffic Knowledge Graph Generation from Scene ImagesYunfei Guo, Fei Yin, Xiao-Hui Li, Xudong Yan et al.ICCV 2023 · 18 citations
- A Siamese Transformer with Hierarchical Refinement for Lane DetectionZinan Lv, Dong Han, Wenzhe Wang, Danny Z. ChenNeurIPS 2024 · 5 citations
Builds on2
- Learning Lightweight Lane Detection CNNs by Self Attention DistillationYuenan Hou, Zheng Ma, Chunxiao Liu, Chen Change LoyICCV 2019 · 666 citations
- Rethinking Semantic Segmentation From a Sequence-to-Sequence Perspective With TransformersSixiao Zheng, Jiachen Lu, Hengshuang Zhao, Xiatian Zhu et al.CVPR 2021
Related papers
- Generating Dynamic Kernels via Transformers for Lane DetectionZiye Chen, Yu Liu, Mingming Gong, Bo Du et al.ICCV 2023 · 25 citations
- RESA: Recurrent Feature-Shift Aggregator for Lane DetectionTu Zheng, Hao Fang, Yi Zhang, Wenjian Tang et al.AAAI 2021 · 348 citations
- GeoFormer: Geometry Point Encoder for 3D Object Detection with Graph-Based TransformerXin Jin, Haisheng Su, Cong Ma, Kai Liu et al.ICCV 2025 · 2 citations
- CrackFormer: Transformer Network for Fine-Grained Crack DetectionHuajun Liu, Xiangyu Miao, Christoph Mertz, Chengzhong Xu et al.ICCV 2021 · 195 citations
- Keep Your Eyes on the Lane: Real-Time Attention-Guided Lane DetectionLucas Tabelini Torres, Rodrigo Ferreira Berriel, Thiago M. Paixão, Claudine Badue et al.CVPR 2021
