MOVIS: Enhancing Multi-Object Novel View Synthesis for Indoor Scenes
Ruijie Lu, Yixin Chen, Junfeng Ni, Baoxiong Jia, Yu Liu, Diwen Wan, Gang Zeng, Siyuan Huang
Abstract
Left 180° Down 5° Left 180° Down 5° Left 180° Down 5° Right 145° Up 10° Right 80° Left 50° Right 100° Up 5° Right 100° Up 5° Right 100° Up 5°O bjaverse SUNRGB-D 3D-FRONT NVS Cross View Matching Ours Zero-1-to-3 Ground Truth Figure 1 . Novel view synthesis and cross-view image matching. The first row shows that MOVIS generalizes to different datasets on novel view synthesis (NVS). We also show visualizations of cross-view consistency compared with Zero-1-to-3 [8] and ground truth by applying image-matching. MOVIS can match a significantly greater number of points, closely aligned with the ground truth.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 20d602de-fb8f-40ed-9c91-babfea5cc4b2Cited by top-tier papers9
- HoloScene: Simulation-Ready Interactive 3D Worlds from a Single VideoHongchi Xia, Chih-Hao Lin, Hao-Yu Hsu, Quentin Leboutet et al.NeurIPS 2025 · 18 citations
- Trace3D: Consistent Segmentation Lifting via Gaussian Instance TracingHongyu Shen, Junfeng Ni, Yixin Chen, Weishuo Li et al.ICCV 2025 · 5 citations
- TACO: Taming Diffusion for In-the-Wild Video Amodal CompletionRuijie Lu, Yixin Chen, Yu Liu, Jiaxiang Tang et al.ICCV 2025 · 3 citations
- Lifting Unlabeled Internet-level Data for 3D Scene UnderstandingYixin Chen, Yaowei Zhang, Huangyue Yu, Junchao He et al.CVPR 2026 · 1 citation
- Masked Point-Entity Contrast for Open-Vocabulary 3D Scene UnderstandingYan Wang, Baoxiong Jia, Ziyu Zhu, Siyuan HuangCVPR 2025
Builds on12
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov et al.ICCV 2023 · 1,662 citations
- One-2-3-45: Any Single Image to 3D Mesh in 45 Seconds without Per-Shape OptimizationMinghua Liu, Chao Xu, Haian Jin, Linghao Chen et al.NeurIPS 2023 · 755 citations
Related papers
- Free3D: Consistent Novel View Synthesis Without 3D RepresentationChuanxia Zheng, Andrea VedaldiCVPR 2024 · 28 citations
- Zero-Shot Novel View and Depth Synthesis with Multi-View Geometric DiffusionVitor Guizilini, Muhammad Zubair Irshad, Dian Chen, Greg Shakhnarovich et al.CVPR 2025
- FreeVS: Generative View Synthesis on Free Driving TrajectoryQitai Wang, Lue Fan, Yuqi Wang, Yuntao Chen et al.ICLR 2025
- Rotationally-Consistent Novel View Synthesis for HumansYoungjoong Kwon, Stefano Petrangeli, Dahun Kim, Haoliang Wang et al.ACM MM 2020 · 5 citations
- MVD-Fusion: Single-view 3D via Depth-consistent Multi-view GenerationHanzhe Hu, Zhizhuo Zhou, Varun Jampani, Shubham TulsianiCVPR 2024
