SimVS: Simulating World Inconsistencies for Robust View Synthesis
Alex Trevithick, Roni Paiss, Philipp Henzler, Dor Verbin, Rundi Wu, Hadi Alzayer, Ruiqi Gao, Ben Poole, Jonathan T. Barron, Aleksander Holynski, Ravi Ramamoorthi, Pratul P. Srinivasan
摘要
Novel-view synthesis techniques achieve impressive results for static scenes but struggle when faced with the inconsistencies inherent to casual capture settings: varying illumination, scene motion, and other unintended effects that are difficult to model explicitly. We present an approach for leveraging generative video models to simulate the inconsistencies in the world that can occur during capture. We use this process, along with existing multi-view datasets, to create synthetic data for training a multi-view harmonization network that is able to reconcile inconsistent observations into a consistent 3D scene. We demonstrate that our world-simulation strategy significantly outperforms traditional augmentation methods in handling real-world scene variations, thereby enabling highly accurate static 3D reconstructions in the presence of a variety of challenging inconsistencies. Project
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- PPISP: Physically-Plausible Compensation and Control of Photometric Variations in Radiance Field ReconstructionIsaac Deutsch, Nicolas Moënne-Loccoz, Gavriel State, Žan GojčičCVPR 2026 · 被引用 7 次
- LuxRemix: Lighting Decomposition and Remixing for Indoor ScenesRuofan Liang, Norman Müller, Ethan Weber, Duncan Zauss 等CVPR 2026 · 被引用 7 次
- Coupled Diffusion Sampling for Training-Free Multi-View Image EditingHadi Alzayer, Yunzhi Zhang, Chen Geng, Jia-Bin Huang 等CVPR 2026 · 被引用 6 次
- UniVerse: Unleashing the Scene Prior of Video Diffusion Models for Robust Radiance Field ReconstructionJin Cao, Hongrui Wu, Ziyong Feng, Hujun Bao 等ICCV 2025 · 被引用 1 次
- CAT4D: Create Anything in 4D with Multi-View Video Diffusion ModelsRundi Wu, Ruiqi Gao, Ben Poole, Alex Trevithick 等CVPR 2025
它引用的顶会 Paper39
- Depth Anything V2Lihe Yang, Bingyi Kang, Zilong Huang, Zhen Zhao 等NeurIPS 2024 · 被引用 2,305 次
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov 等ICCV 2023 · 被引用 1,662 次
- MVSNeRF: Fast Generalizable Radiance Field Reconstruction from Multi-View StereoAnpei Chen, Zexiang Xu, Fuqiang Zhao, Xiaoshuai Zhang 等ICCV 2021 · 被引用 1,024 次
- LRM: Large Reconstruction Model for Single Image to 3DYicong Hong, Kai Zhang, Jiuxiang Gu, Sai Bi 等ICLR 2024 · 被引用 813 次
- Zip-NeRF: Anti-Aliased Grid-Based Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan 等ICCV 2023 · 被引用 799 次
相关 Paper
- Vivid4D: Improving 4D Reconstruction from Monocular Video by Video InpaintingJiaxin Huang, Sheng Miao, Bangbang Yang, Yuewen Ma 等ICCV 2025 · 被引用 2 次
- Deep 3D Mask Volume for View Synthesis of Dynamic ScenesKai-En Lin, Lei Xiao, Feng Liu, Guowei Yang 等ICCV 2021 · 被引用 42 次
- CHROMA: Consistent Harmonization of Multi-View Appearance via Bilateral Grid PredictionJisu Shin, Richard Shaw, Seunghyun Shin, Zhensong Zhang 等ICLR 2026 · 被引用 4 次
- AccidentalGS: 3D Gaussian Splatting from Accidental Camera MotionMao Mao, Xujie Shen, Guyuan Chen, Boming Zhao 等ICCV 2025 · 被引用 1 次
- SparseCam4D: Spatio-Temporally Consistent 4D Reconstruction from Sparse CamerasWeihong Pan, Xiaoyu Zhang, Zhuang Zhang, Zhichao Ye 等CVPR 2026
