Image-guided Neural Object Rendering
Justus Thies, Michael Zollhöfer, Christian Theobalt, Marc Stamminger, Matthias Nießner
Abstract
We propose a learned image-guided rendering technique that combines the benefits of image-based rendering and GAN-based image synthesis. The goal of our method is to generate photo-realistic re-renderings of reconstructed objects for virtual and augmented reality applications (e.g., virtual showrooms, virtual tours & sightseeing, the digital inspection of historical artifacts). A core component of our work is the handling of view-dependent effects. Specifically, we directly train an object-specific deep neural network to synthesize the view-dependent appearance of an object. As input data we are using an RGB video of the object. This video is used to reconstruct a proxy geometry of the object via multi-view stereo. Based on this 3D proxy, the appearance of a captured view can be warped into a new target view as in classical image-based rendering. This warping assumes diffuse surfaces, in case of view-dependent effects, such as specular highlights, it leads to artifacts. To this end, we propose EffectsNet, a deep neural network that predicts view-dependent effects. Based on these estimations, we are able to convert observed images to diffuse images. These diffuse images can be projected into other views. In the target view, our pipeline reinserts the new view-dependent effects. To composite multiple reprojected images to a final output, we learn a composition network that outputs photo-realistic results. Using this image-guided approach, the network does not have to allocate capacity on "remembering" object appearance, instead it learns how to combine the appearance of captured images. We demonstrate the effectiveness of our approach both qualitatively and quantitatively on synthetic as well as on real data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers14
- GRAM: Generative Radiance Manifolds for 3D-Aware Image GenerationYu Deng, Jiaolong Yang, Jianfeng Xiang, Xin TongCVPR 2022 · 189 citations
- ADOP: approximate differentiable one-pixel point renderingDarius Rückert, Linus Franke, Marc StammingerSIGGRAPH 2022 · 127 citations
- NeRF-SR: High Quality Neural Radiance Fields using SupersamplingChen Wang, Xian Wu, Yuan-Chen Guo, Song-Hai Zhang et al.ACM MM 2022 · 115 citations
- Editable free-viewpoint video using a layered neural representationJiakai Zhang, Xinhang Liu, Xinyi Ye, Fuqiang Zhao et al.SIGGRAPH 2021 · 80 citations
- Equivariant Neural RenderingEmilien Dupont, Miguel Bautista Martin, Alex Colburn, Aditya Sankar et al.ICML 2020 · 69 citations
Builds on1
Related papers
- Novel View Synthesis with View-Dependent Effects from a Single ImageJuan Luis Gonzalez Bello, Munchurl KimCVPR 2024
- Single Image Reflection Removal With Physically-Based Training ImagesSoomin Kim, Yuchi Huo, Sung-Eui YoonCVPR 2020
- Holo-Relighting: Controllable Volumetric Portrait Relighting from a Single ImageYiqun Mei, Yu Zeng, He Zhang, Zhixin Shu et al.CVPR 2024 · 11 citations
- Neural Bokeh: Learning Lens Blur for Computational Videography and Out-of-Focus Mixed RealityDavid Mandl, Shohei Mori, Peter Mohr, Yifan Peng et al.IEEE VR 2024 · 5 citations
- PX-NET: Simple and Efficient Pixel-Wise Training of Photometric Stereo NetworksFotios Logothetis, Ignas Budvytis, Roberto Mecca, Roberto CipollaICCV 2021 · 60 citations
