Structured Bird's-Eye-View Traffic Scene Understanding from Onboard Images
Yigit Baran Can, Alexander Liniger, Danda Pani Paudel, Luc Van Gool
摘要
Autonomous navigation requires structured representation of the road network and instance-wise identification of the other traffic agents. Since the traffic scene is defined on the ground plane, this corresponds to scene understanding in the bird’s-eye-view (BEV). However, the onboard cameras of autonomous cars are customarily mounted horizontally for a better view of the surrounding, making this task very challenging. In this work, we study the problem of extracting a directed graph representing the local road network in BEV coordinates, from a single onboard camera image. Moreover, we show that the method can be extended to detect dynamic objects on the BEV plane. The semantics, locations, and orientations of the detected objects together with the road graph facilitates a comprehensive understanding of the scene. Such understanding becomes fundamental for the downstream tasks, such as path planning and navigation. We validate our approach against powerful baselines and show that our network achieves superior performance. We also demonstrate the effects of various design choices through ablation studies. Code: https://github.com/ybarancan/STSU
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper34
- VectorMapNet: End-to-end Vectorized HD Map LearningYicheng Liu, Tianyuan Yuan, Yue Wang, Yilun Wang 等ICML 2023 · 被引用 332 次
- PolarFormer: Multi-Camera 3D Object Detection with Polar TransformerYanqin Jiang, Li Zhang, Zhenwei Miao, Xiatian Zhu 等AAAI 2023 · 被引用 240 次
- MapTR: Structured Modeling and Learning for Online Vectorized HD Map ConstructionBencheng Liao, Shaoyu Chen, Xinggang Wang, Tianheng Cheng 等ICLR 2023 · 被引用 69 次
- LaneSegNet: Map Learning with Lane Segment Perception for Autonomous DrivingTianyu Li, Peijin Jia, Bangjun Wang, Li Chen 等ICLR 2024 · 被引用 69 次
- TopoMLP: A Simple yet Strong Pipeline for Driving Topology ReasoningDongming Wu, Jiahao Chang, Fan Jia, Yingfei Liu 等ICLR 2024 · 被引用 46 次
它引用的顶会 Paper5
- Learning Lightweight Lane Detection CNNs by Self Attention DistillationYuenan Hou, Zheng Ma, Chunxiao Liu, Chen Change LoyICCV 2019 · 被引用 666 次
- DAGMapper: Learning to Map by Discovering Lane TopologyNamdar Homayounfar, Justin Liang, Wei-Chiu Ma, Jack Fan 等ICCV 2019 · 被引用 104 次
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora 等CVPR 2020
- Predicting Semantic Map Representations From Images Using Pyramid Occupancy NetworksThomas Roddick, Roberto CipollaCVPR 2020
- MP3: A Unified Model To Map, Perceive, Predict and PlanSergio Casas, Abbas Sadat, Raquel UrtasunCVPR 2021
相关 Paper
- Topology Preserving Local Road Network Estimation from Single Onboard Camera ImageYigit Baran Can, Alexander Liniger, Danda Pani Paudel, Luc Van GoolCVPR 2022 · 被引用 37 次
- Improving Online Lane Graph Extraction by Object-Lane ClusteringYigit Baran Can, Alexander Liniger, Danda Pani Paudel, Luc Van GoolICCV 2023 · 被引用 11 次
- 'The Pedestrian next to the Lamppost" Adaptive Object Graphs for Better Instantaneous MappingAvishkar Saha, Oscar Mendez, Chris Russell, Richard BowdenCVPR 2022 · 被引用 8 次
- BEV-Guided Multi-Modality Fusion for Driving PerceptionYunze Man, Liang-Yan Gui, Yu-Xiong WangCVPR 2023
- BEV-SAN: Accurate BEV 3D Object Detection via Slice Attention NetworksXiaowei Chi, Jiaming Liu, Ming Lu, Rongyu Zhang 等CVPR 2023
