GeoSim: Realistic Video Simulation via Geometry-Aware Composition for Self-Driving
Yun Chen, Frieda Rong, Shivam Duggal, Shenlong Wang, Xinchen Yan, Sivabalan Manivasagam, Shangjie Xue, Ersin Yumer, Raquel Urtasun
Abstract
Scalable sensor simulation is an important yet challenging open problem for safety-critical domains such as selfdriving. Current works in image simulation either fail to be photorealistic or do not model the 3D environment and the dynamic objects within, losing high-level control and physical realism. In this paper, we present GeoSim, a geometry-aware image composition process which synthesizes novel urban driving scenarios by augmenting existing images with dynamic objects extracted from other scenes and rendered at novel poses. Towards this goal, we first build a diverse bank of 3D objects with both realistic geometry and appearance from sensor data. During simulation, we perform a novel geometry-aware simulationby-composition procedure which 1) proposes plausible and realistic object placements into a given scene, 2) renders novel views of dynamic objects from the asset bank, and 3) composes and blends the rendered image segments. The resulting synthetic images are realistic, traffic-aware, and geometrically consistent, allowing our approach to scale to complex use cases. We demonstrate two such important applications: long-range realistic video simulation across multiple camera sensors, and synthetic data generation for data augmentation on downstream segmentation tasks. Please check https://tmux.top/publication/geosim/ for high-resolution video results.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers23
- Block-NeRF: Scalable Large Scene Neural View SynthesisMatthew Tancik, Vincent Casser, Xinchen Yan, Sabeek Pradhan et al.CVPR 2022 · 702 citations
- OBJECT 3DIT: Language-guided 3D-aware Image EditingOscar Michel, Anand Bhattad, Eli VanderBilt, Ranjay Krishna et al.NeurIPS 2023 · 79 citations
- Real-Time Neural Rasterization for Large ScenesJeffrey Yunfan Liu, Yun Chen, Ze Yang, Jingkang Wang et al.ICCV 2023 · 46 citations
- Dynamic Mesh-Aware Radiance FieldsYi-Ling Qiao, Alexander Gao, Yiran Xu, Yue Feng et al.ICCV 2023 · 32 citations
- Learning Human Dynamics in Autonomous Driving ScenariosJingbo Wang, Ye Yuan, Zhengyi Luo, Kevin Xie et al.ICCV 2023 · 31 citations
Builds on13
- Free-Form Image Inpainting With Gated ConvolutionJiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen et al.ICCV 2019 · 1,990 citations
- Habitat: A Platform for Embodied AI ResearchManolis Savva, Jitendra Malik, Devi Parikh, Dhruv Batra et al.ICCV 2019 · 1,863 citations
- Everybody Dance NowCaroline Chan, Shiry Ginosar, Tinghui Zhou, Alexei A. EfrosICCV 2019 · 840 citations
- Soft Rasterizer: A Differentiable Renderer for Image-Based 3D ReasoningShichen Liu, Weikai Chen, Tianye Li, Hao LiICCV 2019 · 789 citations
- Learning Joint 2D-3D Representations for Depth CompletionYun Chen, Bin Yang, Ming Liang, Raquel UrtasunICCV 2019 · 190 citations
Related papers
- LiDARsim: Realistic LiDAR Simulation by Leveraging the Real WorldSivabalan Manivasagam, Shenlong Wang, Kelvin Wong, Wenyuan Zeng et al.CVPR 2020
- SurfelGAN: Synthesizing Realistic Sensor Data for Autonomous DrivingZhenpei Yang, Yuning Chai, Dragomir Anguelov, Yin Zhou et al.CVPR 2020
- Neural Lighting Simulation for Urban ScenesAva Pun, Gary Sun, Jingkang Wang, Yun Chen et al.NeurIPS 2023 · 11 citations
- AdvSim: Generating Safety-Critical Scenarios for Self-Driving VehiclesJingkang Wang, Ava Pun, James Tu, Sivabalan Manivasagam et al.CVPR 2021
- SceneGen: Learning To Generate Realistic Traffic ScenesShuhan Tan, Kelvin Wong, Shenlong Wang, Sivabalan Manivasagam et al.CVPR 2021
