Free3D: Consistent Novel View Synthesis Without 3D Representation
Chuanxia Zheng, Andrea Vedaldi
摘要
We introduce Free3D, a simple accurate method for monocular open-set novel view synthesis (NVS). Similar to Zero-1-to-S, we start from a pre-trained 2D image generator for generalization, and fine-tune it for NVS. Compared to other works that took a similar approach, we obtain significant improvements without resorting to an explicit 3D representation, which is slow and memory-consuming, and without training an additional network for 3D reconstruction. Our key contribution is to improve the way the target camera pose is encoded in the network, which we do by introducing a new ray conditioning normalization (RCN) layer. The latter injects pose information in the underlying 2D image generator by telling each pixel its viewing direction. We further improve multi-view consistency by using light-weight multi-view attention layers and by sharing generation noise between the different views. We train Free3D on the Objaverse dataset and demonstrate excellent generalization to new categories in new datasets, including OmniObject3D and GSO. The project page is available at https://chuanxiar.com/free3d/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper41
- Splatter Image: Ultra-Fast Single-View 3D ReconstructionStanislaw Szymanowicz, Christian Rupprecht, Andrea VedaldiCVPR 2024 · 被引用 132 次
- MVSplat360: Feed-Forward 360 Scene Synthesis from Sparse ViewsYuedong Chen, Chuanxia Zheng, Haofei Xu, Bohan Zhuang 等NeurIPS 2024 · 被引用 126 次
- Vivid-ZOO: Multi-View Video Generation with Diffusion ModelBing Li, Cheng Zheng, Wenxuan Zhu, Jinjie Mai 等NeurIPS 2024 · 被引用 48 次
- Stable Virtual Camera: Generative View Synthesis with Diffusion ModelsJensen Zhou, Hang Gao, Vikram Voleti, Aaryaman Vasishta 等ICCV 2025 · 被引用 25 次
- Atlas3D: Physically Constrained Self-Supporting Text-to-3D for Simulation and FabricationYunuo Chen, Tianyi Xie, Zeshun Zong, Xuan Li 等NeurIPS 2024 · 被引用 24 次
它引用的顶会 Paper37
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
相关 Paper
- FreeVS: Generative View Synthesis on Free Driving TrajectoryQitai Wang, Lue Fan, Yuqi Wang, Yuntao Chen 等ICLR 2025
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov 等ICCV 2023 · 被引用 1,662 次
- NViST: In the Wild New View Synthesis from a Single Image with TransformersWonbong Jang, Lourdes AgapitoCVPR 2024
- MOVIS: Enhancing Multi-Object Novel View Synthesis for Indoor ScenesRuijie Lu, Yixin Chen, Junfeng Ni, Baoxiong Jia 等CVPR 2025
- Rayzer: a Self-Supervised Large View Synthesis ModelHanwen Jiang, Hao Tan, Peng Wang, Hai Jin 等ICCV 2025 · 被引用 12 次
