SimVS: Simulating World Inconsistencies for Robust View Synthesis
Alex Trevithick, Roni Paiss, Philipp Henzler, Dor Verbin, Rundi Wu, Hadi Alzayer, Ruiqi Gao, Ben Poole, Jonathan T. Barron, Aleksander Holynski, Ravi Ramamoorthi, Pratul P. Srinivasan
Abstract
Novel-view synthesis techniques achieve impressive results for static scenes but struggle when faced with the inconsistencies inherent to casual capture settings: varying illumination, scene motion, and other unintended effects that are difficult to model explicitly. We present an approach for leveraging generative video models to simulate the inconsistencies in the world that can occur during capture. We use this process, along with existing multi-view datasets, to create synthetic data for training a multi-view harmonization network that is able to reconcile inconsistent observations into a consistent 3D scene. We demonstrate that our world-simulation strategy significantly outperforms traditional augmentation methods in handling real-world scene variations, thereby enabling highly accurate static 3D reconstructions in the presence of a variety of challenging inconsistencies. Project
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2fab6c79-e69d-462e-9558-4af4c8027a6bCited by top-tier papers5
- PPISP: Physically-Plausible Compensation and Control of Photometric Variations in Radiance Field ReconstructionIsaac Deutsch, Nicolas Moënne-Loccoz, Gavriel State, Žan GojčičCVPR 2026 · 7 citations
- LuxRemix: Lighting Decomposition and Remixing for Indoor ScenesRuofan Liang, Norman Müller, Ethan Weber, Duncan Zauss et al.CVPR 2026 · 7 citations
- Coupled Diffusion Sampling for Training-Free Multi-View Image EditingHadi Alzayer, Yunzhi Zhang, Chen Geng, Jia-Bin Huang et al.CVPR 2026 · 6 citations
- UniVerse: Unleashing the Scene Prior of Video Diffusion Models for Robust Radiance Field ReconstructionJin Cao, Hongrui Wu, Ziyong Feng, Hujun Bao et al.ICCV 2025 · 1 citation
- CAT4D: Create Anything in 4D with Multi-View Video Diffusion ModelsRundi Wu, Ruiqi Gao, Ben Poole, Alex Trevithick et al.CVPR 2025
Builds on39
- Depth Anything V2Lihe Yang, Bingyi Kang, Zilong Huang, Zhen Zhao et al.NeurIPS 2024 · 2,305 citations
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov et al.ICCV 2023 · 1,662 citations
- MVSNeRF: Fast Generalizable Radiance Field Reconstruction from Multi-View StereoAnpei Chen, Zexiang Xu, Fuqiang Zhao, Xiaoshuai Zhang et al.ICCV 2021 · 1,024 citations
- LRM: Large Reconstruction Model for Single Image to 3DYicong Hong, Kai Zhang, Jiuxiang Gu, Sai Bi et al.ICLR 2024 · 813 citations
- Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan et al.ICCV 2023 · 799 citations
Related papers
- Vivid4D: Improving 4D Reconstruction from Monocular Video by Video InpaintingJiaxin Huang, Sheng Miao, Bangbang Yang, Yuewen Ma et al.ICCV 2025 · 2 citations
- Deep 3D Mask Volume for View Synthesis of Dynamic ScenesKai-En Lin, Lei Xiao, Feng Liu, Guowei Yang et al.ICCV 2021 · 42 citations
- CHROMA: Consistent Harmonization of Multi-View Appearance via Bilateral Grid PredictionJisu Shin, Richard Shaw, Seunghyun Shin, Zhensong Zhang et al.ICLR 2026 · 4 citations
- AccidentalGS: 3D Gaussian Splatting from Accidental Camera MotionMao Mao, Xujie Shen, Guyuan Chen, Boming Zhao et al.ICCV 2025 · 1 citation
- SparseCam4D: Spatio-Temporally Consistent 4D Reconstruction from Sparse CamerasWeihong Pan, Xiaoyu Zhang, Zhuang Zhang, Zhichao Ye et al.CVPR 2026
