Poirot: Automatic Root Cause Analysis of Safety Violations in ADS Simulation Testing via Hypothetical Reasoning
You Lu, Dingji Wang, Kun Zhang, Bihuan Chen, Jiyan Zhang, Xin Peng
摘要
With the rapid development of autonomous driving systems (ADSs), it has become critical to ensure their operational safety, leading to the widespread adoption of simulation testing. While existing scenario-based simulation testing approaches have demonstrated effectiveness in detecting safety violations, they often fall short in providing insight into the underlying causes of these violations, which is an essential capability for improving the safety and reliability of ADSs. To address this limitation, we propose a two-phase novel framework, Poirot, for root cause analysis in simulation testing via hypothetical reasoning. Given a reproducible violation scenario, in the module-level analysis phase, Poirot replays the violation scenario and identifies the faulty module by iteratively replacing an actual module with an idealized module and checking whether the violation persists. In the component-level analysis phase, depending on the identified faulty module, Poirot further applies either hypothetical reasoning with a suspicion-guided search strategy or causal analysis to narrow the fault space and pinpoint the faulty component. We evaluate Poirot with two ADSs, e.g., Apollo and Autoware, on a comprehensive benchmark that includes a total of 80 real and injected faults along with their triggering scenarios. Compared with the state-of-the-art root cause analysis approaches, e.g., ACAV and Rocas, Poirot improves the module-level accuracy by 187.29% on average, and identifies the faulty components at a finer granularity, achieving component-level accuracy of 90.62%. Our ablation study shows that our suspicion-guided search strategy in Poirot efficiently reduces the exploration of the fault space by 58.77%, leading to a 65.41% reduction in the time for fault localization. Finally, applied to two scenario-based simulation testing methods, i.e., AvFuzzer and MoDitector, Poirot attributes 425 violation scenarios to 8 faults, cutting debugging time by 96.89% compared to manual analysis in practice.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- MoDitector: Module-Directed Testing for Autonomous Driving SystemsRenzhi Wang, Mingfei Cheng, Xiaofei Xie, Yuan Zhou 等ISSTA 2025 · 被引用 3 次
- Interventional Root Cause Analysis of Failures in Multi-Sensor Fusion Perception SystemsShuguang Wang, Qian Zhou, Kui Wu, Jinghuai Deng 等NDSS 2025
- DiaVio: LLM-Empowered Diagnosis of Safety Violations in ADS Simulation TestingYou Lu, Yifan Tian, Yuyang Bi, Bihuan Chen 等ISSTA 2024 · 被引用 9 次
- ACAV: A Framework for Automatic Causality Analysis in Autonomous Vehicle Accident RecordingsHuijia Sun, Christopher M. Poskitt, Yang Sun, Jun Sun 等ICSE 2024 · 被引用 11 次
- VioHawk: Detecting Traffic Violations of Autonomous Driving Systems through Criticality-Guided Simulation TestingZhongrui Li, Jiarun Dai, Zongan Huang, Nianhao You 等ISSTA 2024 · 被引用 7 次
