A Large-Scale Outdoor Multi-modal Dataset and Benchmark for Novel View Synthesis and Implicit Scene Reconstruction
Chongshan Lu, Fukun Yin, Xin Chen, Wen Liu, Tao Chen, Gang Yu, Jiayuan Fan
Abstract
Neural Radiance Fields (NeRF) [24] has achieved impressive results in single object scene reconstruction and novel view synthesis, as demonstrated on many single modality and single object focused indoor scene datasets like DTU [14], BMVS [42], and NeRF Synthetic [24]. However, the study of NeRF on large-scale outdoor scene reconstruction is still limited, as there is no unified outdoor scene dataset for large-scale NeRF evaluation due to expensive data acquisition and calibration costs. In this work, we propose a large-scale outdoor multi-modal dataset, OMMO dataset, containing complex objects and scenes with calibrated images, point clouds and prompt annotations. A new benchmark for several outdoor NeRF-based tasks is established, such as novel view synthesis, diverse 3D representation, and multi-modal NeRF. To create the dataset, we capture and collect a large number of real fly-view videos and select high-quality and high-resolution clips from them. Then we design a quality review module to refine images, remove low-quality frames and fail-to-calibrate scenes through a learning-based automatic evaluation plus manual review. Finally, volunteers are employed to label and review the prompt annotation for each scene and keyframe. Compared with existing NeRF datasets, our dataset contains abundant real-world urban and natural scenes with various scales, camera trajectories, and lighting conditions. Experiments show that our dataset can benchmark most state-of-the-art NeRF methods on different tasks. The dataset can be found at the following link: https://ommo.luchongshan.com/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1c197926-323f-4455-83de-72c10a91cf70Cited by top-tier papers15
- 3D Gaussian Splatting as Markov Chain Monte CarloShakiba Kheradmand, Daniel Rebain, Gopal Sharma, Weiwei Sun et al.NeurIPS 2024 · 285 citations
- MatrixCity: A Large-scale City Dataset for City-scale Neural Rendering and BeyondYixuan Li, Lihan Jiang, Linning Xu, Yuanbo Xiangli et al.ICCV 2023 · 185 citations
- ClimateNeRF: Extreme Weather Synthesis in Neural Radiance FieldYuan Li, Zhi-Hao Lin, David A. Forsyth, Jia-Bin Huang et al.ICCV 2023 · 44 citations
- ChatCam: Empowering Camera Control through Conversational AIXinhang Liu, Yu-Wing Tai, Chi-Keung TangNeurIPS 2024 · 19 citations
- PDF: Point Diffusion Implicit Function for Large-scale Scene Neural RepresentationYuhan Ding, Fukun Yin, Jiayuan Fan, Hui Li et al.NeurIPS 2023 · 7 citations
Builds on19
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman et al.ICCV 2021 · 2,700 citations
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt et al.NeurIPS 2021 · 2,500 citations
- Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan et al.CVPR 2022 · 1,603 citations
Related papers
- NeRF in the Wild: Neural Radiance Fields for Unconstrained Photo CollectionsRicardo Martin-Brualla, Noha Radwan, Mehdi S. M. Sajjadi, Jonathan T. Barron et al.CVPR 2021
- ReLight My NeRF: A Dataset for Novel View Synthesis and Relighting of Real World ObjectsMarco Toschi, Riccardo De Matteo, Riccardo Spezialetti, Daniele De Gregorio et al.CVPR 2023
- PKU-DyMVHumans: A Multi-View Video Benchmark for High-Fidelity Dynamic Human ModelingXiaoyun Zheng, Liwei Liao, Xufeng Li, Jianbo Jiao et al.CVPR 2024 · 7 citations
- Few-Shot Neural Radiance Fields under Unconstrained IlluminationSeokYeong Lee, Junyong Choi, Seungryong Kim, Ig-Jae Kim et al.AAAI 2024 · 11 citations
- S-NeRF: Neural Radiance Fields for Street ViewsZiyang Xie, Junge Zhang, Wenye Li, Feihu Zhang et al.ICLR 2023 · 13 citations
