AR-NeRF: Unsupervised Learning of Depth and Defocus Effects from Natural Images with Aperture Rendering Neural Radiance Fields
Takuhiro Kaneko
Abstract
Fully unsupervised 3D representation learning has gained attention owing to its advantages in data collection. A successful approach involves a viewpoint-aware approach that learns an image distribution based on generative models (e.g., generative adversarial networks (GANs)) while generating various view images based on 3D-aware models (e.g., neural radiance fields (NeRFs)). However, they require images with various views for training, and consequently, their application to datasets with few or limited viewpoints remains a challenge. As a complementary approach, an aperture rendering GAN (AR-GAN) that employs a defocus cue was proposed. However, an AR-GAN is a CNN-based model and represents a defocus independently from a viewpoint change despite its high correlation, which is one of the reasons for its performance. As an alternative to an AR-GAN, we propose an aperture rendering NeRF (AR-NeRF), which can utilize viewpoint and defocus cues in a unified manner by representing both factors in a common ray-tracing framework. Moreover, to learn defocus-aware and defocus-independent representations in a disentangled manner, we propose aperture randomized training, for which we learn to generate images while randomizing the aperture size and latent codes independently. During our experiments, we applied AR-NeRF to various natural image datasets, including flower, bird, and face images, the results of which demonstrate the utility of AR-NeRF for un-supervised learning of the depth and defocus effects.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1d4b91bf-e67a-4fd7-81d3-94522e94a601Cited by top-tier papers6
- NaviNeRF: NeRF-based 3D Representation Disentanglement by Latent Semantic NavigationBaao Xie, Bohan Li, Zequn Zhang, Junting Dong et al.ICCV 2023 · 12 citations
- MIMO-NeRF: Fast Neural Rendering with Multi-input Multi-output Neural Radiance FieldsTakuhiro KanekoICCV 2023 · 7 citations
- Strata-NeRF : Neural Radiance Fields for Stratified ScenesAnkit Dhiman, R. Srinath, Harsh Rangwani, Rishubh Parihar et al.ICCV 2023 · 5 citations
- Improving Physics-Augmented Continuum Neural Radiance Field-Based Geometry-Agnostic System Identification with Lagrangian Particle OptimizationTakuhiro KanekoCVPR 2024
- Inverting the Imaging Process by Learning an Implicit Camera ModelXin Huang, Qi Zhang, Ying Feng, Hongdong Li et al.CVPR 2023
Builds on32
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil et al.NeurIPS 2020 · 4,036 citations
- FaceForensics++: Learning to Detect Manipulated Facial ImagesAndreas Rössler, Davide Cozzolino, Luisa Verdoliva, Christian Riess et al.ICCV 2019 · 2,966 citations
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen et al.NeurIPS 2021 · 2,126 citations
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- PlenOctrees for Real-time Rendering of Neural Radiance FieldsAlex Yu, Ruilong Li, Matthew Tancik, Hao Li et al.ICCV 2021 · 1,284 citations
Related papers
- DoF-NeRF: Depth-of-Field Meets Neural Radiance FieldsZijin Wu, Xingyi Li, Juewen Peng, Hao Lu et al.ACM MM 2022 · 34 citations
- Pix2NeRF: Unsupervised Conditional -GAN for Single Image to Neural Radiance Fields TranslationShengqu Cai, Anton Obukhov, Dengxin Dai, Luc Van GoolCVPR 2022 · 70 citations
- ABLE-NeRF: Attention-Based Rendering with Learnable Embeddings for Neural Radiance FieldZhe Jun Tang, Tat-Jen Cham, Haiyu ZhaoCVPR 2023
- 3D-aware Image Synthesis via Learning Structural and Textural RepresentationsYinghao Xu, Sida Peng, Ceyuan Yang, Yujun Shen et al.CVPR 2022 · 88 citations
- RegNeRF: Regularizing Neural Radiance Fields for View Synthesis from Sparse InputsMichael Niemeyer, Jonathan T. Barron, Ben Mildenhall, Mehdi S. M. Sajjadi et al.CVPR 2022 · 513 citations
