ViewFusion: Towards Multi-View Consistency via Interpolated Denoising
Xianghui Yang, Yan Zuo, Sameera Ramasinghe, Loris Bazzani, Gil Avraham, Anton van den Hengel
摘要
Novel-view synthesis through diffusion models has demonstrated remarkable potential for generating diverse and high-quality images. Yet, the independent process of image generation in these prevailing methods leads to challenges in maintaining multiple view consistency. To address this, we introduce ViewFusion, a novel, training-free algorithm that can be seamlessly integrated into existing pre-trained diffusion models. Our approach adopts an auto-regressive method that implicitly leverages previously generated views as context for next view generation, ensuring robust multi-view consistency during the novel-view generation process. Through a diffusion process that fuses known-view information via interpolated denoising, our framework successfully extends single-view conditioned models to work in multiple-view conditional settings without any additional fine-tuning. Extensive experimental results demonstrate the effectiveness of View Fusion in generating consistent and detailed novel views.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Zero-to-Hero: Enhancing Zero-Shot Novel View Synthesis via Attention Map FilteringIdo Sobol, Chenfeng Xu, Or LitanyNeurIPS 2024 · 被引用 10 次
- MaterialMVP: Illumination-Invariant Material Generation via Multi-View PBR DiffusionZebin He, Mingxin Yang, Shuhui Yang, Yixuan Tang 等ICCV 2025 · 被引用 3 次
- Benchmarking and Learning Multi-Dimensional Quality Evaluator for Text-To-3D GenerationYujie Zhang, Bingyang Cui, Qi Yang, Zhu Li 等ICCV 2025 · 被引用 3 次
- MVGBench: A Comprehensive Benchmark for Multi-View Generation ModelsXianghui Xie, Jan Eric Lenssen, Gerard Pons-MollICCV 2025 · 被引用 2 次
- Consistency of Compositional Generalization Across Multiple LevelsChuanhao Li, Zhen Li, Chenchen Jing, Xiaomeng Fan 等AAAI 2025 · 被引用 1 次
它引用的顶会 Paper47
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt 等NeurIPS 2021 · 被引用 2,500 次
相关 Paper
- WAVE: Warp-Based View Guidance for Consistent Novel View Synthesis Using a Single ImageJiwoo Park, Tae Eun Choi, Youngjun Jun, Seong Jae HwangICCV 2025
- ViVid-1-to-3: Novel View Synthesis with Video Diffusion ModelsJeong-gi Kwak, Erqun Dong, Yuhe Jin, Hanseok Ko 等CVPR 2024
- DreamComposer: Controllable 3D Object Generation via Multi-View ConditionsYunhan Yang, Yukun Huang, Xiaoyang Wu, Yuan-Chen Guo 等CVPR 2024 · 被引用 3 次
- TexPainter: Generative Mesh Texturing with Multi-view ConsistencyHongkun Zhang, Zherong Pan, Congyi Zhang, Lifeng Zhu 等SIGGRAPH 2024 · 被引用 18 次
- SemanticNVS: Improving Semantic Scene Understanding in Generative Novel View SynthesisXinya Chen, Christopher Wewer, Jiahao Xie, Xinting Hu 等ICML 2026
