Visual Traffic Knowledge Graph Generation from Scene Images
Yunfei Guo, Fei Yin, Xiao-Hui Li, Xudong Yan, Tao Xue, Shuqi Mei, Cheng-Lin Liu
Abstract
Although previous works on traffic scene understanding have achieved great success, most of them stop at a lowlevel perception stage, such as road segmentation and lane detection, and few concern high-level understanding. In this paper, we present Visual Traffic Knowledge Graph Generation (VTKGG), a new task for in-depth traffic scene understanding that tries to extract multiple kinds of information and integrate them into a knowledge graph. To achieve this goal, we first introduce a large dataset named CASIA-Tencent Road Scene dataset (RS10K) with comprehensive annotations to support related research. Secondly, we propose a novel traffic scene parsing architecture containing a Hierarchical Graph ATtention network (HGAT) to analyze the heterogeneous elements and their complicated relations in traffic scene images. By hierarchizing the heterogeneous graph and equipping it with cross-level links, our approach exploits the correlation among various elements completely and acquires accurate relations. The experimental results show that our method can effectively generate visual traffic knowledge graphs and achieve state-of-the-art performance. The dataset RS10K is available at http: //www.nlpr.ia.ac.cn/pal/RS10K.html .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e0d77943-beb4-4c71-9482-38e10b0257f9Cited by top-tier papers6
- Aligning Vision to Language: Annotation-Free Multimodal Knowledge Graph Construction for Enhanced LLMs ReasoningJunming Liu, Siyuan Meng, Yanting Gao, Song Mao et al.ICCV 2025 · 34 citations
- Drive-R1: Bridging Reasoning and Planning in VLMs for Autonomous Driving with Reinforcement LearningYue Li, Meng Tian, Dechang Zhu, Jiangtong Zhu et al.AAAI 2026 · 27 citations
- Fine-Grained Evaluation of Large Vision-Language Models in Autonomous DrivingYue Li, Meng Tian, Zhenyu Lin, Jiangtong Zhu et al.ICCV 2025 · 4 citations
- Multi-modal Traffic Scenario Generation for Autonomous Driving System TestingZhi Tu, Liangkun Niu, Wei Fan, Tianyi ZhangFSE 2025 · 1 citation
- Driving by the Rules: A Benchmark for Integrating Traffic Sign Regulations into Vectorized HD MapXinyuan Chang, Maixuan Xue, Xinran Liu, Zheng Pan et al.CVPR 2025
Builds on8
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 6,042 citations
- CLRNet: Cross Layer Refinement Network for Lane DetectionTu Zheng, Yifei Huang, Yang Liu, Wenjian Tang et al.CVPR 2022 · 280 citations
- Rethinking Efficient Lane Detection via Curve ModelingZhengyang Feng, Shaohua Guo, Xin Tan, Ke Xu et al.CVPR 2022 · 204 citations
- Structured Bird's-Eye-View Traffic Scene Understanding from Onboard ImagesYigit Baran Can, Alexander Liniger, Danda Pani Paudel, Luc Van GoolICCV 2021 · 147 citations
- Laneformer: Object-Aware Row-Column Transformers for Lane DetectionJianhua Han, Xiajun Deng, Xinyue Cai, Zhen Yang et al.AAAI 2022 · 79 citations
Related papers
- Learning to Understand Traffic SignsYunfei Guo, Wei Feng, Fei Yin, Tao Xue et al.ACM MM 2021 · 13 citations
- Traffic Scene Parsing Through the TSP6K DatasetPeng-Tao Jiang, Yuqi Yang, Yang Cao, Qibin Hou et al.CVPR 2024
- SUTD-TrafficQA: A Question Answering Benchmark and an Efficient Network for Video Reasoning Over Traffic EventsLi Xu, He Huang, Jun LiuCVPR 2021
- Progressive Graph Attention Network for Video Question AnsweringLiang Peng, Shuangji Yang, Yi Bin, Guoqing WangACM MM 2021 · 47 citations
- HL-Net: Heterophily Learning Network for Scene Graph GenerationXin Lin, Changxing Ding, Yibing Zhan, Zijian Li et al.CVPR 2022 · 51 citations
