Fine-Grained Representation for Lane Topology Reasoning
Guoqing Xu, Yiheng Li, Yang Yang
Abstract
Precise modeling of lane topology is essential for autonomous driving, as it directly impacts navigation and control decisions. Existing methods typically represent each lane with a single query and infer topological connectivity based on the similarity between lane queries. However, this kind of design struggles to accurately model complex lane structures, leading to unreliable topology prediction. In this view, we propose a Fine-Grained lane topology reasoning framework (TopoFG). It divides the procedure from bird’s-eye-view (BEV) features to topology prediction via fine-grained queries into three phases, i.e., Hierarchical Prior Extractor (HPE), Region-Focused Decoder (RFD), and Robust Boundary-Point Topology Reasoning (RBTR). Specifically, HPE extracts global spatial priors from the BEV mask and local sequential priors from in-lane keypoint sequences to guide subsequent fine-grained query modeling. RFD constructs fine-grained queries by integrating the spatial and sequential priors. It then samples reference points in RoI regions of the mask and applies cross-attention with BEV features to refine the query representations of each lane. RBTR models lane connectivity based on boundary-point query features and further employs a topological denoising strategy to reduce matching ambiguity. By integrating spatial and sequential priors into fine-grained queries and applying a denoising strategy to boundary-point topology reasoning, our method precisely models complex lane structures and delivers trustworthy topology predictions. Extensive experiments on the OpenLane-V2 benchmark demonstrate that TopoFG achieves new state-of-the-art performance, with an OLS of 48.0% on subset_A and 45.4% on subset_B.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on22
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- Vision Transformer with Deformable AttentionZhuofan Xia, Xuran Pan, Shiji Song, Li Erran Li et al.CVPR 2022 · 835 citations
- VectorMapNet: End-to-end Vectorized HD Map LearningYicheng Liu, Tianyuan Yuan, Yue Wang, Yilun Wang et al.ICML 2023 · 332 citations
- 3D-LaneNet: End-to-End 3D Multiple Lane DetectionNoa Garnett, Rafi Cohen, Tomer Pe'er, Roee Lahav et al.ICCV 2019 · 232 citations
- Structured Bird's-Eye-View Traffic Scene Understanding from Onboard ImagesYigit Baran Can, Alexander Liniger, Danda Pani Paudel, Luc Van GoolICCV 2021 · 147 citations
Related papers
- TopoPoint: Enhance Topology Reasoning via Endpoint Detection in Autonomous DrivingYanping Fu, Xinyuan Liu, Tianyu Li, Yike Ma et al.NeurIPS 2025 · 10 citations
- TopoHR: Hierarchical Centerline Representation for Cyclic Topology Reasoning in Driving Scenes with Point-to-Instance RelationsYifeng Bai, Zhirong Chen, Bo Song, Erkang Cheng et al.CVPR 2026 · 1 citation
- Topo2Seq: Enhanced Topology Reasoning via Topology Sequence LearningYiming Yang, Yueru Luo, Bingkun He, Erlong Li et al.AAAI 2025 · 8 citations
- Geometry-Guided Representations for Coherent Lane and Traffic Topology Reasoning in Driving ScenesYueru Luo, Changqing Zhou, Yiming Yang, Erlong Li et al.KDD 2026 · 2 citations
- Driving Scene Understanding with Traffic Scene-Assisted Topology Graph TransformerFu Rong, Wenjin Peng, Meng Lan, Qian Zhang et al.ACM MM 2024 · 5 citations
