How Does Simulation-Based Testing for Self-Driving Cars Match Human Perception?
Christian Birchler, Tanzil Kombarabettu Mohammed, Pooja Rani, Teodora Nechita, Timo Kehrer, Sebastiano Panichella
Abstract
Software metrics such as coverage and mutation scores have been extensively explored for the automated quality assessment of test suites. While traditional tools rely on such quantifiable software metrics, the field of self-driving cars (SDCs) has primarily focused on simulation-based test case generation using quality metrics such as the out-of-bound (OOB) parameter to determine if a test case fails or passes. However, it remains unclear to what extent this quality metric aligns with the human perception of the safety and realism of SDCs, which are critical aspects in assessing SDC behavior. To address this gap, we conducted an empirical study involving 50 participants to investigate the factors that determine how humans perceive SDC test cases as safe, unsafe, realistic, or unrealistic. To this aim, we developed a framework leveraging virtual reality (VR) technologies, called SDC-Alabaster, to immerse the study participants into the virtual environment of SDC simulators. Our findings indicate that the human assessment of the safety and realism of failing and passing test cases can vary based on different factors, such as the test's complexity and the possibility of interacting with the SDC. Especially for the assessment of realism, the participants' age as a confounding factor leads to a different perception. This study highlights the need for more research on SDC simulation testing quality metrics and the importance of human perception in evaluating SDC behavior.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0354bcf4-9f09-4de5-8ae3-03d23f52d4fcBuilds on8
- DeepBillboard: systematic physical-world testing of autonomous driving systemsHusheng Zhou, Wei Li, Zelun Kong, Junfeng Guo et al.ICSE 2020 · 150 citations
- Misbehaviour prediction for autonomous driving systemsAndrea Stocco, Michael Weiss, Marco Calzana, Paolo TonellaICSE 2020 · 138 citations
- A comprehensive study of autonomous vehicle bugsJoshua Garcia, Yang Feng, Junjie Shen, Sumaya Almanee et al.ICSE 2020 · 127 citations
- WalkingVibe: Reducing Virtual Reality Sickness and Improving Realism while Walking in VR using Unobtrusive Head-mounted Vibrotactile FeedbackYi-Hao Peng, Carolyn Yu, Shi-Hong Liu, Chung-Wei Wang et al.CHI 2020 · 75 citations
- An exploratory study of autopilot software bugs in unmanned aerial vehiclesDinghua Wang, Shuqing Li, Guanping Xiao, Yepang Liu et al.FSE 2021 · 60 citations
Related papers
- Toward Immersive Self-Driving Simulations: Reports from a User Study across Six PlatformsDohyeon Yeo, Gwangbin Kim, Seungjun KimCHI 2020 · 49 citations
- Virtual Morality: Using Virtual Reality to Study Moral Behavior in Extreme Accident SituationsGiulia Benvegnù, Patrik Pluchino, Luciano GamberiniIEEE VR 2021 · 19 citations
- Generating Realistic and Diverse Tests for LiDAR-Based Perception SystemsGarrett Christian, Trey Woodlief, Sebastian G. ElbaumICSE 2023 · 11 citations
- Building Critical Testing Scenarios for Autonomous Driving from Real AccidentsXudong Zhang, Yan CaiISSTA 2023 · 30 citations
- A Multi-Modality Evaluation of the Reality Gap in Autonomous Driving SystemsStefano Carlo Lambertenghi, Mirena Flores Valdez, Andrea StoccoASE 2025 · 1 citation
