NeRF-Supervised Deep Stereo
Fabio Tosi, Alessio Tonioni, Daniele De Gregorio, Matteo Poggi
Abstract
We introduce a novel framework for training deep stereo networks effortlessly and without any ground-truth. By leveraging state-of-the-art neural rendering solutions, we generate stereo training data from image sequences collected with a single handheld camera. On top of them, a NeRF-supervised training procedure is carried out, from which we exploit rendered stereo triplets to compensate for occlusions and depth maps as proxy labels. This results in stereo networks capable of predicting sharp and detailed disparity maps. Experimental results show that models trained under this regime yield a 30-40% improvement over existing self-supervised methods on the challenging Middle-bury dataset, filling the gap to supervised models and, most times, outperforming them at zero-shot generalization.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8df399f5-b4c5-4724-b492-77bb5b207e2eCited by top-tier papers16
- UC-NERF: Neural Radiance Field for Under-Calibrated Multi-View Cameras in Autonomous DrivingKai Cheng, Xiaoxiao Long, Wei Yin, Jin Wang et al.ICLR 2024 · 26 citations
- Active Stereo Without Pattern ProjectorLuca Bartolomei, Matteo Poggi, Fabio Tosi, Andrea Conti et al.ICCV 2023 · 12 citations
- Robust Synthetic-to-Real Transfer for Stereo MatchingJiawei Zhang, Jiahe Li, Lei Huang, Xiaohan Yu et al.CVPR 2024 · 12 citations
- Federated Online Adaptation for Deep StereoMatteo Poggi, Fabio TosiCVPR 2024 · 11 citations
- SeaScan: An Energy-Efficient Underwater Camera for Wireless 3D Color ImagingNazish Naeem, Jack Rademacher, Ritik Patnaik, Tara Boroushaki et al.MobiCom 2024 · 7 citations
Builds on33
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman et al.ICCV 2021 · 2,700 citations
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 2,416 citations
- Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan et al.CVPR 2022 · 1,603 citations
- Plenoxels: Radiance Fields without Neural NetworksSara Fridovich-Keil, Alex Yu, Matthew Tancik, Qinhong Chen et al.CVPR 2022 · 1,237 citations
Related papers
- GS-ASM: 2DGS-Supervised Active Stereo MatchingZhengling Wu, Rongfeng Lu, Quan Chen, Longjian Zeng et al.CVPR 2026
- ZeroStereo: Zero-Shot Stereo Matching from Single ImagesXianqi Wang, Hao Yang, Gangwei Xu, Junda Cheng et al.ICCV 2025 · 1 citation
- Generalizable Novel-View Synthesis Using a Stereo CameraHaechan Lee, Wonjoon Jin, Seung-Hwan Baek, Sunghyun ChoCVPR 2024
- SceneRF: Self-Supervised Monocular 3D Scene Reconstruction with Radiance FieldsAnh-Quan Cao, Raoul de CharetteICCV 2023 · 73 citations
- Learning to Render Novel Views from Wide-Baseline Stereo PairsYilun Du, Cameron Smith, Ayush Tewari, Vincent SitzmannCVPR 2023
