StylizedNeRF: Consistent 3D Scene Stylization as Stylized NeRF via 2D-3D Mutual Learning
Yihua Huang, Yue He, Yu-Jie Yuan, Yu-Kun Lai, Lin Gao
Abstract
3D scene stylization aims at generating stylized images of the scene from arbitrary novel views following a given set of style examples, while ensuring consistency when rendered from different views. Directly applying methods for image or video stylization to 3D scenes cannot achieve such consistency. Thanks to recently proposed neural radiance fields (NeRF), we are able to represent a 3D scene in a consistent way. Consistent 3D scene stylization can be effectively achieved by stylizing the corresponding NeRF. However, there is a significant domain gap between style examples which are 2D images and NeRF which is an implicit volumetric representation. To address this problem, we propose a novel mutual learning framework for 3D scene stylization that combines a 2D image stylization network and NeRF to fuse the stylization ability of 2D stylization network with the 3D consistency of NeRF. We first pre-train a standard NeRF of the 3D scene to be stylized and replace its color prediction module with a style network to obtain a stylized NeRF. It is followed by distilling the prior knowledge of spatial consistency from NeRF to the 2D stylization network through an introduced consistency loss. We also introduce a mimic loss to supervise the mutual learning of the NeRF style module and fine-tune the 2D stylization decoder. In order to further make our model handle ambiguities of 2D stylization results, we introduce learnable latent codes that obey the probability distributions conditioned on the style. They are attached to training samples as conditional inputs to better learn the style module in our novel stylized NeRF. Experimental results demonstrate that our method is superior to existing approaches in both visual quality and long-range consistency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 46fab328-16c0-47a3-87be-81c49314ece3Cited by top-tier papers48
- Instruct-NeRF2NeRF: Editing 3D Scenes with InstructionsAyaan Haque, Matthew Tancik, Alexei A. Efros, Aleksander Holynski et al.ICCV 2023 · 544 citations
- AvatarCraft: Transforming Text into Neural Human Avatars with Parameterized Shape and Pose ControlRuixiang Jiang, Can Wang, Jingbo Zhang, Menglei Chai et al.ICCV 2023 · 103 citations
- ViCA-NeRF: View-Consistency-Aware 3D Editing of Neural Radiance FieldsJiahua Dong, Yu-Xiong WangNeurIPS 2023 · 97 citations
- FocalDreamer: Text-Driven 3D Editing via Focal-Fusion AssemblyYuhan Li, Yishun Dou, Yue Shi, Yu Lei et al.AAAI 2024 · 91 citations
- GaussianEditor: Editing 3D Gaussians Delicately with Text InstructionsJunjie Wang, Jiemin Fang, Xiaopeng Zhang, Lingxi Xie et al.CVPR 2024 · 65 citations
Builds on15
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil et al.NeurIPS 2020 · 4,036 citations
- Neural Sparse Voxel FieldsLingjie Liu, Jiatao Gu, Kyaw Zaw Lin, Tat-Seng Chua et al.NeurIPS 2020 · 1,535 citations
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun et al.NeurIPS 2020 · 1,010 citations
- NeRD: Neural Reflectance Decomposition from Image CollectionsMark Boss, Raphael Braun, Varun Jampani, Jonathan T. Barron et al.ICCV 2021 · 608 citations
- Photorealistic Style Transfer via Wavelet TransformsJaejun Yoo, Youngjung Uh, Sanghyuk Chun, Byeongkyu Kang et al.ICCV 2019 · 412 citations
Related papers
- SNeRF: stylized neural implicit representations for 3D scenesThu Nguyen-Phuoc, Feng Liu, Lei XiaoSIGGRAPH 2022 · 99 citations
- Locally Stylized Neural Radiance FieldsHong-Wing Pang, Binh-Son Hua, Sai-Kit YeungICCV 2023 · 18 citations
- StyleNeRF: A Style-based 3D Aware Generator for High-resolution Image SynthesisJiatao Gu, Lingjie Liu, Peng Wang, Christian TheobaltICLR 2022 · 622 citations
- Transforming Radiance Field with Lipschitz Network for Photorealistic 3D Scene StylizationZicheng Zhang, Yinglu Liu, Congying Han, Yingwei Pan et al.CVPR 2023
- NeRF Analogies: Example-Based Visual Attribute Transfer for NeRFsMichael Fischer, Zhengqin Li, Thu Nguyen-Phuoc, Aljaz Bozic et al.CVPR 2024
