ViVid-1-to-3: Novel View Synthesis with Video Diffusion Models
Jeong-gi Kwak, Erqun Dong, Yuhe Jin, Hanseok Ko, Shweta Mahajan, Kwang Moo Yi
Abstract
https://ubc-vision.github.io/vivid123/ View-conditioned Diffusion model (Zero-1-to-3 XL) ViVid Guidance Input image Target camera trajectory Target camera trajectory Figure 1 . Teaser -we present a strikingly simple training-free method to make already available novel-view synthesis diffusion models more consistent both in terms of the desired viewing angle and the content-combining it with video diffusion. As shown in the example, the results of our method are more consistent with the input images and correspond more to the target views.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 95ba6f3f-6ba1-4d8d-972b-bc356c6db47aCited by top-tier papers36
- CAT3D: Create Anything in 3D with Multi-View Diffusion ModelsRuiqi Gao, Aleksander Holynski, Philipp Henzler, Arthur Brussee et al.NeurIPS 2024 · 490 citations
- IM-3D: Iterative Multiview Diffusion and Reconstruction for High-Quality 3D GenerationLuke Melas-Kyriazi, Iro Laina, Christian Rupprecht, Natalia Neverova et al.ICML 2024 · 92 citations
- VideoGPA: Distilling Geometry Priors for 3D-Consistent Video GenerationHongyang Du, Hongyang Du, Xiaoyan Cong, Runhao Li et al.ICML 2026 · 16 citations
- Zero-to-Hero: Enhancing Zero-Shot Novel View Synthesis via Attention Map FilteringIdo Sobol, Chenfeng Xu, Or LitanyNeurIPS 2024 · 10 citations
- Edit360: 2D Image Edits to 3D Assets From Any AngleJunchao Huang, Xinting Hu, Shaoshuai Shi, Zhuotao Tian et al.ICCV 2025 · 2 citations
Builds on39
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language ModelsJunnan Li, Dongxu Li, Silvio Savarese, Steven C. H. HoiICML 2023 · 7,873 citations
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan et al.NeurIPS 2022 · 2,948 citations
Related papers
- ViewFusion: Towards Multi-View Consistency via Interpolated DenoisingXianghui Yang, Yan Zuo, Sameera Ramasinghe, Loris Bazzani et al.CVPR 2024 · 5 citations
- WAVE: Warp-Based View Guidance for Consistent Novel View Synthesis Using a Single ImageJiwoo Park, Tae Eun Choi, Youngjun Jun, Seong Jae HwangICCV 2025
- Consistent View Synthesis with Pose-Guided Diffusion ModelsHung-Yu Tseng, Qinbo Li, Changil Kim, Suhib Alsisan et al.CVPR 2023
- Vistadream: Sampling Multiview Consistent Images for Single-View Scene ReconstructionHaiping Wang, Yuan Liu, Ziwei Liu, Wenping Wang et al.ICCV 2025 · 8 citations
- Look Beyond: Two-Stage Scene View Generation via Panorama and Video DiffusionXueyang Kang, Zhengkang Xiang, Zezheng Zhang, Kourosh KhoshelhamACM MM 2025
