MapTR: Structured Modeling and Learning for Online Vectorized HD Map Construction
Bencheng Liao, Shaoyu Chen, Xinggang Wang, Tianheng Cheng, Qian Zhang, Wenyu Liu, Chang Huang
摘要
High-definition (HD) map provides abundant and precise environmental information of the driving scene, serving as a fundamental and indispensable component for planning in autonomous driving system. We present MapTR, a structured end-to-end Transformer for efficient online vectorized HD map construction. We propose a unified permutation-equivalent modeling approach, i.e., modeling map element as a point set with a group of equivalent permutations, which accurately describes the shape of map element and stabilizes the learning process. We design a hierarchical query embedding scheme to flexibly encode structured map information and perform hierarchical bipartite matching for map element learning. MapTR achieves the best performance and efficiency with only camera input among existing vectorized map construction approaches on nuScenes dataset. In particular, MapTR-nano runs at real-time inference speed ( FPS) on RTX 3090, faster than the existing state-of-the-art camera-based method while achieving higher mAP. Even compared with the existing state-of-the-art multi-modality method, MapTR-nano achieves higher mAP, and MapTR-tiny achieves higher mAP and faster inference speed. Abundant qualitative results show that MapTR maintains stable and robust map construction quality in complex and various driving scenes. MapTR is of great application value in autonomous driving. Code and more demos are available at https://github.com/hustvl/MapTR.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper66
- Vista: A Generalizable Driving World Model with High Fidelity and Versatile ControllabilityShenyuan Gao, Jiazhi Yang, Li Chen, Kashyap Chitta 等NeurIPS 2024 · 被引用 403 次
- PivotNet: Vectorized Pivot Learning for End-to-end HD Map ConstructionWenjie Ding, Limeng Qiao, Xi Qiu, Chi ZhangICCV 2023 · 被引用 119 次
- RAD: Training an End-to-End Driving Policy via Large-Scale 3DGS-based Reinforcement LearningHao Gao, Shaoyu Chen, Bo Jiang, Bencheng Liao 等NeurIPS 2025 · 被引用 92 次
- Online Map Vectorization for Autonomous Driving: A Rasterization PerspectiveGongjie Zhang, Jiahao Lin, Shuang Wu, Yilin Song 等NeurIPS 2023 · 被引用 78 次
- LaneSegNet: Map Learning with Lane Segment Perception for Autonomous DrivingTianyu Li, Peijin Jia, Bangjun Wang, Li Chen 等ICLR 2024 · 被引用 69 次
它引用的顶会 Paper18
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- BEVDepth: Acquisition of Reliable Depth for Multi-View 3D Object DetectionYinhao Li, Zheng Ge, Guanyi Yu, Jinrong Yang 等AAAI 2023 · 被引用 954 次
- You Only Look at One Sequence: Rethinking Transformer in Vision through Object DetectionYuxin Fang, Bencheng Liao, Xinggang Wang, Jiemin Fang 等NeurIPS 2021 · 被引用 430 次
- Cross-view Transformers for real-time Map-view Semantic SegmentationBrady Zhou, Philipp KrähenbühlCVPR 2022 · 被引用 279 次
相关 Paper
- VectorMapNet: End-to-end Vectorized HD Map LearningYicheng Liu, Tianyuan Yuan, Yue Wang, Yilun Wang 等ICML 2023 · 被引用 332 次
- Learning Global Representation from Queries for Vectorized HD Map ConstructionShoumeng Qiu, Xinrun Li, Yang Long, Xiangyang Xue 等ICML 2026 · 被引用 1 次
- End-to-End Vectorized HD-map Construction with Piecewise Bézier CurveLimeng Qiao, Wenjie Ding, Xi Qiu, Chi ZhangCVPR 2023
- InteractionMap: Improving Online Vectorized HDMap Construction with InteractionKuang Wu, Chuan Yang, Zhanbin LiCVPR 2025
- Compact HD Map Construction via Douglas-Peucker Point TransformerRuixin Liu, Zejian YuanAAAI 2024 · 被引用 6 次
