Panoptic Compositional Feature Field for Editable Scene Rendering with Network-Inferred Labels via Metric Learning
Xinhua Cheng, Yanmin Wu, Mengxi Jia, Qian Wang, Jian Zhang
摘要
Despite neural implicit representations demonstrating impressive high-quality view synthesis capacity, decomposing such representations into objects for instance-level editing is still challenging. Recent works learn objectcompositional representations supervised by ground truth instance annotations and produce promising scene editing results. However, ground truth annotations are manually labeled and expensive in practice, which limits their usage in real-world scenes. In this work, we attempt to learn an object-compositional neural implicit representation for editable scene rendering by leveraging labels inferred from the off-the-shelf 2D panoptic segmentation networks instead of the ground truth annotations. We propose a novel framework named Panoptic Compositional Feature Field (PCFF), which introduces an instance quadruplet metric learning to build a discriminating panoptic feature space for reliable scene editing. In addition, we propose semanticrelated strategies to further exploit the correlations between semantic and appearance attributes for achieving better rendering results. Experiments on multiple scene datasets including ScanNet, Replica, and ToyDesk demonstrate that our proposed method achieves superior performance for novel view synthesis and produces convincing real-world scene editing results.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Progressive3D: Progressively Local Editing for Text-to-3D Content Creation with Complex Semantic PromptsXinhua Cheng, Tianyu Yang, Jianan Wang, Yu Li 等ICLR 2024 · 被引用 58 次
- OmniSeg3D: Omniversal 3D Segmentation via Hierarchical Contrastive LearningHaiyang Ying, Yixuan Yin, Jinzhi Zhang, Fan Wang 等CVPR 2024 · 被引用 32 次
- Aerial Lifting: Neural Urban Semantic and Building Instance Lifting from Aerial ImageryYuqi Zhang, Guanying Chen, Jiaxing Chen, Shuguang CuiCVPR 2024 · 被引用 4 次
- Retri3D: 3D Neural Graphics Representation RetrievalYushi Guan, Daniel Kwan, Jean Sebastien Dandurand, Xi Yan 等ICLR 2025
它引用的顶会 Paper30
- Mip-NeRF: A Multiscale Representation for Anti-Aliasing Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Matthew Tancik, Peter Hedman 等ICCV 2021 · 被引用 2,700 次
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt 等NeurIPS 2021 · 被引用 2,500 次
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 被引用 2,196 次
- Neural Sparse Voxel FieldsLingjie Liu, Jiatao Gu, Kyaw Zaw Lin, Tat-Seng Chua 等NeurIPS 2020 · 被引用 1,535 次
- PlenOctrees for Real-time Rendering of Neural Radiance FieldsAlex Yu, Ruilong Li, Matthew Tancik, Hao Li 等ICCV 2021 · 被引用 1,284 次
相关 Paper
- Panoptic Neural Fields: A Semantic Object-Aware Neural Scene RepresentationAbhijit Kundu, Kyle Genova, Xiaoqi Yin, Alireza Fathi 等CVPR 2022 · 被引用 204 次
- Learning Object-Compositional Neural Radiance Field for Editable Scene RenderingBangbang Yang, Yinda Zhang, Yinghao Xu, Yijin Li 等ICCV 2021 · 被引用 305 次
- AssetField: Assets Mining and Reconfiguration in Ground Feature Plane RepresentationYuanbo Xiangli, Linning Xu, Xingang Pan, Nanxuan Zhao 等ICCV 2023 · 被引用 13 次
- Learning Unified Decompositional and Compositional NeRF for Editable Novel View SynthesisYuxin Wang, Wayne Wu, Dan XuICCV 2023 · 被引用 18 次
- RICO: Regularizing the Unobservable for Indoor Compositional ReconstructionZizhang Li, Xiaoyang Lyu, Yuanyuan Ding, Mengmeng Wang 等ICCV 2023 · 被引用 17 次
