SMD: Multi-view Safety-Critical Driving Video Generation in the Real-world Domain
Jiawei Zhou, Linye Lyu, Zhuotao Tian, Cheng Zhuo, YU LI
Abstract
Safety-critical scenarios are rare yet pivotal for evaluating and enhancing the robustness of autonomous driving systems. While existing methods generate safety-critical driving trajectories, simulations, or single-view videos, they fall short of meeting the demands of advanced end-to-end autonomous systems (E2E AD), which require real-world, multi-view video data. To bridge this gap, we introduce SafeMVDrive, the first framework designed to generate high-quality, safety-critical, multi-view driving videos grounded in real-world domains. SafeMV-Drive strategically integrates a safety-critical trajectory generator with an advanced multi-view video generator. To tackle the challenges inherent in this integration, we first enhance scene understanding ability of the trajectory generator by incorporating visual context -which is previously unavailable to such generator -and leveraging a GRPO-finetuned vision-language model to achieve more realistic and context-aware trajectory generation. Second, recognizing that existing multi-view video generators struggle to render realistic collision events, we introduce a twostage, controllable trajectory generation mechanism that produces collision-evasion trajectories, ensuring both video quality and safety-critical fidelity. Finally, we employ a diffusion-based multi-view video generator to synthesize high-quality safety-critical driving videos from the generated trajectories. Experiments conducted on an E2E AD planner demonstrate a significant increase in collision rate when tested with our generated data, validating the effectiveness of SafeMVDrive in stress-testing planning modules. Our code, examples, and datasets are publicly available at: https://zhoujiawei3.github.io/SafeMVDrive/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1a4680cf-ace5-48ca-812e-201492a71384Cited by top-tier papers1
Ask how each one uses itBuilds on14
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Vista: A Generalizable Driving World Model with High Fidelity and Versatile ControllabilityShenyuan Gao, Jiazhi Yang, Li Chen, Kashyap Chitta et al.NeurIPS 2024 · 403 citations
- Large Language Models Are Reasoning TeachersNamgyu Ho, Laura Schmid, Se-Young YunACL 2023 · 102 citations
- DiffScene: Diffusion-Based Safety-Critical Scenario Generation for Autonomous VehiclesChejian Xu, Aleksandr Petiushko, Ding Zhao, Bo LiAAAI 2025 · 90 citations
- ChatScene: Knowledge-Enabled Safety-Critical Scenario Generation for Autonomous VehiclesJiawei Zhang, Chejian Xu, Bo LiCVPR 2024 · 50 citations
Related papers
- InstaDrive: Instance-Aware Driving World Models for Realistic and Consistent Video GenerationZhuoran Yang, Xi Guo, Chenjing Ding, Chiyu Wang et al.ICCV 2025 · 3 citations
- SGDrive: Scene-to-Goal Hierarchical World Cognition for Autonomous Drivingjingyu li, Junjie Wu, Dongnan Hu, Xiangkai Huang et al.CVPR 2026 · 36 citations
- SafeDrive: Fine-Grained Safety Reasoning for End-to-End Driving in a Sparse WorldJungho Kim, Jiyong Oh, Seunghoon Yu, Hongjae Shin et al.CVPR 2026 · 8 citations
- Driving Into the Future: Multiview Visual Forecasting and Planning with World Model for Autonomous DrivingYuqi Wang, Jiawei He, Lue Fan, Hongxin Li et al.CVPR 2024
- SynAD: Enhancing Real-World End-to-End Autonomous Driving Models through Synthetic Data IntegrationJongsuk Kim, Jaeyoung Lee, Gyojin Han, Dong-Jae Lee et al.ICCV 2025 · 5 citations
