LeGEND: A Top-Down Approach to Scenario Generation of Autonomous Driving Systems Assisted by Large Language Models
Shuncheng Tang, Zhenya Zhang, Jixiang Zhou, Lei Lei, Yuan Zhou, Yinxing Xue
Abstract
Autonomous driving systems (ADS) are safety-critical and require comprehensive testing before their deployment on public roads. While existing testing approaches primarily aim at the criticality of scenarios, they often overlook the diversity of the generated scenarios that is also important to reflect system defects in different aspects. To bridge the gap, we propose LeGEND, that features a top-down fashion of scenario generation: it starts with abstract functional scenarios, and then steps downwards to logical and concrete scenarios, such that scenario diversity can be controlled at the functional level. However, unlike logical scenarios that can be formally described, functional scenarios are often documented in natural languages (e.g., accident reports) and thus cannot be precisely parsed and processed by computers. To tackle that issue, LeGEND leverages the recent advances of large language models (LLMs) to transform textual functional scenarios to formal logical scenarios. To mitigate the distraction of useless information in functional scenario description, we devise a two-phase transformation that features the use of an intermediate language; consequently, we adopt two LLMs in LeGEND, one for extracting information from functional scenarios, the other for converting the extracted information to formal logical scenarios. We experimentally evaluate LeGEND on Apollo, an industry-grade ADS from Baidu. Evaluation results show that LeGEND can effectively identify critical scenarios, and compared to baseline approaches, LeGEND exhibits evident superiority in diversity of generated scenarios. Moreover, we also demonstrate the advantages of our two-phase transformation framework, and the accuracy of the adopted LLMs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 042a6e13-80aa-4d57-b0ab-7afa0ac14ab6Cited by top-tier papers3
- Decictor: Towards Evaluating the Robustness of Decision-Making in Autonomous Driving SystemsMingfei Cheng, Xiaofei Xie, Yuan Zhou, Junjie Wang et al.ICSE 2025 · 4 citations
- Multi-modal Traffic Scenario Generation for Autonomous Driving System TestingZhi Tu, Liangkun Niu, Wei Fan, Tianyi ZhangFSE 2025 · 1 citation
- TrafficAlign: Aligning Large Language Models for Traffic Scenario GenerationZhi Tu, Liangkun Niu, Tianyi ZhangCVPR 2026
Builds on12
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Invisible for both Camera and LiDAR: Security of Multi-Sensor Fusion based Perception in Autonomous Driving Under Physical-World AttacksYulong Cao, Ningfei Wang, Chaowei Xiao, Dawei Yang et al.S&P 2021 · 309 citations
- CodaMosa: Escaping Coverage Plateaus in Test Generation with Pre-trained Large Language ModelsCaroline Lemieux, Jeevana Priya Inala, Shuvendu K. Lahiri, Siddhartha SenICSE 2023 · 221 citations
- Model-based exploration of the frontier of behaviours for deep learning system testingVincenzo Riccio, Paolo TonellaFSE 2020 · 134 citations
- DeepHyperion: exploring the feature space of deep learning-based systems through illumination searchTahereh Zohdinasab, Vincenzo Riccio, Alessio Gambi, Paolo TonellaISSTA 2021 · 76 citations
Related papers
- DiaVio: LLM-Empowered Diagnosis of Safety Violations in ADS Simulation TestingYou Lu, Yifan Tian, Yuyang Bi, Bihuan Chen et al.ISSTA 2024 · 9 citations
- Fixed-Point Guided ADS Scenario Generation via Multi-modal LLM Reasoning and Software TestingXudong Zhang, Shihao Zhu, Yan CaiISSTA 2026
- MOSAT: finding safety violations of autonomous driving systems using multi-objective genetic algorithmHaoxiang Tian, Yan Jiang, Guoquan Wu, Jiren Yan et al.FSE 2022 · 74 citations
- SCTrans: Constructing a Large Public Scenario Dataset for Simulation Testing of Autonomous Driving SystemsJiarun Dai, Bufan Gao, Mingyuan Luo, Zongan Huang et al.ICSE 2024 · 8 citations
- Generating Critical Test Scenarios for Autonomous Driving Systems via Influential Behavior PatternsHaoxiang Tian, Guoquan Wu, Jiren Yan, Yan Jiang et al.ASE 2022 · 24 citations
