SceneRF: Self-Supervised Monocular 3D Scene Reconstruction with Radiance Fields
Anh-Quan Cao, Raoul de Charette
Abstract
3D reconstruction from a single 2D image was extensively covered in the literature but relies on depth supervision at training time, which limits its applicability. To relax the dependence to depth we propose SceneRF, a self-supervised monocular scene reconstruction method using only posed image sequences for training. Fueled by the recent progress in neural radiance fields (NeRF) we optimize a radiance field though with explicit depth optimization and a novel probabilistic sampling strategy to efficiently handle large scenes. At inference, a single input image suffices to hallucinate novel depth views which are fused together to obtain 3D scene reconstruction. Thorough experiments demonstrate that we outperform all baselines for novel depth views synthesis and scene reconstruction, on indoor BundleFusion and outdoor SemanticKITTI. Code is available at https://astra-vision.github.io/SceneRF.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2f3b3325-a27f-46f0-ae1a-e7253af5294fCited by top-tier papers25
- Scene as OccupancyWenwen Tong, Chonghao Sima, Tai Wang, Li Chen et al.ICCV 2023 · 251 citations
- TIP-Editor: An Accurate 3D Editor Following Both Text-Prompts And Image-PromptsJingyu Zhuang, Di Kang, Yan-Pei Cao, Guanbin Li et al.SIGGRAPH 2024 · 59 citations
- MonoNeRD: NeRF-like Representations for Monocular 3D Object DetectionJunkai Xu, Liang Peng, Haoran Chen, Hao Li et al.ICCV 2023 · 54 citations
- DVGT: Driving Visual Geometry TransformerSicheng Zuo, Zixun Xie, Wenzhao Zheng, Shaoqing Xu et al.CVPR 2026 · 23 citations
- QuadricFormer: Scene as Superquadrics for 3D Semantic Occupancy PredictionSicheng Zuo, Wenzhao Zheng, Xiaoyong Han, Longchao Yang et al.NeurIPS 2025 · 23 citations
Builds on32
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel et al.ICCV 2019 · 2,345 citations
- Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan et al.CVPR 2022 · 1,603 citations
- PointFlow: 3D Point Cloud Generation With Continuous Normalizing FlowsGuandao Yang, Xun Huang, Zekun Hao, Ming-Yu Liu et al.ICCV 2019 · 794 citations
- Depth-supervised NeRF: Fewer Views and Faster Training for FreeKangle Deng, Andrew Liu, Jun-Yan Zhu, Deva RamananCVPR 2022 · 756 citations
- Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category ReconstructionJeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone et al.ICCV 2021 · 686 citations
Related papers
- MonoNeRF: Learning Generalizable NeRFs from Monocular Videos without Camera PosesYang Fu, Ishan Misra, Xiaolong WangICML 2023 · 13 citations
- LOLNeRF: Learn from One LookDaniel Rebain, Mark J. Matthews, Kwang Moo Yi, Dmitry Lagun et al.CVPR 2022
- AltNeRF: Learning Robust Neural Radiance Field via Alternating Depth-Pose OptimizationKun Wang, Zhiqiang Yan, Huang Tian, Zhenyu Zhang et al.AAAI 2024 · 6 citations
- Dense Depth Priors for Neural Radiance Fields from Sparse Input ViewsBarbara Roessle, Jonathan T. Barron, Ben Mildenhall, Pratul P. Srinivasan et al.CVPR 2022 · 319 citations
- Shape, Pose, and Appearance from a Single Image via Bootstrapped Radiance Field InversionDario Pavllo, David Joseph Tan, Marie-Julie Rakotosaona, Federico TombariCVPR 2023
