SMD: Multi-view Safety-Critical Driving Video Generation in the Real-world Domain
Jiawei Zhou, Linye Lyu, Zhuotao Tian, Cheng Zhuo, YU LI
摘要
Safety-critical scenarios are rare yet pivotal for evaluating and enhancing the robustness of autonomous driving systems. While existing methods generate safety-critical driving trajectories, simulations, or single-view videos, they fall short of meeting the demands of advanced end-to-end autonomous systems (E2E AD), which require real-world, multi-view video data. To bridge this gap, we introduce SafeMVDrive, the first framework designed to generate high-quality, safety-critical, multi-view driving videos grounded in real-world domains. SafeMV-Drive strategically integrates a safety-critical trajectory generator with an advanced multi-view video generator. To tackle the challenges inherent in this integration, we first enhance scene understanding ability of the trajectory generator by incorporating visual context -which is previously unavailable to such generator -and leveraging a GRPO-finetuned vision-language model to achieve more realistic and context-aware trajectory generation. Second, recognizing that existing multi-view video generators struggle to render realistic collision events, we introduce a twostage, controllable trajectory generation mechanism that produces collision-evasion trajectories, ensuring both video quality and safety-critical fidelity. Finally, we employ a diffusion-based multi-view video generator to synthesize high-quality safety-critical driving videos from the generated trajectories. Experiments conducted on an E2E AD planner demonstrate a significant increase in collision rate when tested with our generated data, validating the effectiveness of SafeMVDrive in stress-testing planning modules. Our code, examples, and datasets are publicly available at: https://zhoujiawei3.github.io/SafeMVDrive/ .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper14
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Vista: A Generalizable Driving World Model with High Fidelity and Versatile ControllabilityShenyuan Gao, Jiazhi Yang, Li Chen, Kashyap Chitta 等NeurIPS 2024 · 被引用 403 次
- Large Language Models Are Reasoning TeachersNamgyu Ho, Laura Schmid, Se-Young YunACL 2023 · 被引用 102 次
- DiffScene: Diffusion-Based Safety-Critical Scenario Generation for Autonomous VehiclesChejian Xu, Aleksandr Petiushko, Ding Zhao, Bo LiAAAI 2025 · 被引用 90 次
- ChatScene: Knowledge-Enabled Safety-Critical Scenario Generation for Autonomous VehiclesJiawei Zhang, Chejian Xu, Bo LiCVPR 2024 · 被引用 50 次
相关 Paper
- InstaDrive: Instance-Aware Driving World Models for Realistic and Consistent Video GenerationZhuoran Yang, Xi Guo, Chenjing Ding, Chiyu Wang 等ICCV 2025 · 被引用 3 次
- SGDrive: Scene-to-Goal Hierarchical World Cognition for Autonomous Drivingjingyu li, Junjie Wu, Dongnan Hu, Xiangkai Huang 等CVPR 2026 · 被引用 36 次
- SafeDrive: Fine-Grained Safety Reasoning for End-to-End Driving in a Sparse WorldJungho Kim, Jiyong Oh, Seunghoon Yu, Hongjae Shin 等CVPR 2026 · 被引用 8 次
- Driving Into the Future: Multiview Visual Forecasting and Planning with World Model for Autonomous DrivingYuqi Wang, Jiawei He, Lue Fan, Hongxin Li 等CVPR 2024
- SynAD: Enhancing Real-World End-to-End Autonomous Driving Models through Synthetic Data IntegrationJongsuk Kim, Jaeyoung Lee, Gyojin Han, Dong-Jae Lee 等ICCV 2025 · 被引用 5 次
