GP-NeRF: Generalized Perception NeRF for Context-Aware 3D Scene Understanding
Hao Li, Dingwen Zhang, Yalun Dai, Nian Liu, Lechao Cheng, Jingfeng Li, Jingdong Wang, Junwei Han
Abstract
Figure 1. Our method, called GP-NeRF, achieves remarkable performance improvements for instance and semantic segmentation in both synthesis [35] and real-world [10] datasets, as shown in the right column of the figure. Here we showcase generalized semantic segmentation, finetuning semantic segmentation, and instance segmentation) with their corresponding reconstruction results. For the left column, the qualitative results of the visualization are presented, showing the effectiveness of our method for simultaneous segmentation and reconstruction. What's more, we visualize our rendered features via PCA in the novel view, demonstrating our method possesses the capability to produce semantic-aware features that can distinguish between different classes and objects.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext df477384-5905-4d52-ad17-18224e404944Cited by top-tier papers7
- STRIDER: Navigation via Instruction-Aligned Structural Decision Space OptimizationDiqi He, Xuehao Gao, Hao Li, Junwei Han et al.NeurIPS 2025 · 8 citations
- LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video DiffusionFangfu Liu, Hao Li, Jiawei Chi, Hanyang Wang et al.ICCV 2025 · 7 citations
- LoopGaussian: Creating 3D Cinemagraph with Multi-view Images via Eulerian Motion FieldJiyang Li, Lechao Cheng, Zhangye Wang, Tingting Mu et al.ACM MM 2024 · 4 citations
- ClaraVid: A Holistic Scene Reconstruction Benchmark from Aerial Perspective with Delentropy-Based Complexity ProfilingRadu Beche, Sergiu NedevschiICCV 2025 · 4 citations
- CityGS-: A Scalable Architecture for Efficient and Geometrically Accurate Large-Scale Scene ReconstructionYuanyuan Gao, Hao Li, Jiaqi Chen, Zhengyu Zou et al.ICCV 2025 · 2 citations
Builds on32
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman et al.ICCV 2021 · 2,700 citations
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 2,196 citations
- Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan et al.CVPR 2022 · 1,603 citations
- Nerfies: Deformable Neural Radiance FieldsKeunhong Park, Utkarsh Sinha, Jonathan T. Barron, Sofien Bouaziz et al.ICCV 2021 · 1,442 citations
Related papers
- GSNeRF: Generalizable Semantic Neural Radiance Fields with Enhanced 3D Scene UnderstandingZi-Ting Chou, Sheng-Yu Huang, I-Jieh Liu, Yu-Chiang Frank WangCVPR 2024
- Gear-NeRF: Free-Viewpoint Rendering and Tracking with Motion-Aware Spatio-Temporal SamplingXinhang Liu, Yu-Wing Tai, Chi-Keung Tang, Pedro Miraldo et al.CVPR 2024
- Instance Neural Radiance FieldYichen Liu, Benran Hu, Junkai Huang, Yu-Wing Tai et al.ICCV 2023 · 49 citations
- VisFusion: Visibility-Aware Online 3D Scene Reconstruction from VideosHuiyu Gao, Wei Mao, Miaomiao LiuCVPR 2023
- Taming Generative Diffusion Model for Task-Oriented Infrared ImagingTengyu Ma, Zhilong Dai, Yubo Diao, Guanming An et al.CVPR 2026
