Rethinking End-to-End 2D to 3D Scene Segmentation in Gaussian Splatting
Runsong Zhu, Shi Qiu, Zhengzhe Liu, Ka-Hei Hui, Qianyi Wu, Pheng-Ann Heng, Chi-Wing Fu
摘要
Lifting multi-view 2D instance segmentation to a radiance field has proven effective to enhance 3D understanding. Existing works rely on direct matching for end-to-end lifting, yielding inferior results, or employ a two-stage solution constrained by complex pre-or post-processing. In this work, we design Unified-Lift, a new end-to-end objectaware lifting approach that aims for high-quality 3D segmentation based on our object-aware 3D Gaussian representation. To start, we augment each Gaussian point with a Gaussian-level feature learned using a contrastive loss to encode instance information. Importantly, we introduce a learnable object-level codebook to account for individual objects in the scene for an explicit object-level understanding and associate the encoded object-level features with the Gaussian-level point features for segmentation predictions. While promising, achieving effective codebook learning is nontrivial and a naive solution leads to degraded performance. Hence, we formulate the association learning module and the noisy label filtering module for effective and robust codebook learning. We conduct experiments on three benchmarks LERF-Masked, Replica, and Messy Rooms. Both qualitative and quantitative results manifest that our Unified-Lift clearly outperforms existing methods in terms of segmentation quality and time efficiency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- COS3D: Collaborative Open-Vocabulary 3D SegmentationRunsong Zhu, Ka-Hei Hui, Zhengzhe Liu, Qianyi Wu 等NeurIPS 2025 · 被引用 12 次
- EPS3D: End-to-End Feed-Forward 3D Panoptic SegmentationRunsong Zhu, Jiaxin GUO, Xiaoyang Guo, Zhengzhe Liu 等ICML 2026 · 被引用 3 次
- BEA-GS: BEyond RAdiance Supervision in 3DGS for Precise Object ExtractionAlessio Mazzucchelli, Maria Naranjo-Almeida, Jorge Bustos-Sanchez, Mariella Dimiccoli 等CVPR 2026
- B-Seg: Camera-Free, Training-Free 3DGS Segmentation via Analytic EIG and Beta-Bernoulli Bayesian UpdatesHiromichi Kamata, Samuel Arthur Munro, Fuminori HommaCVPR 2026
- IGFuse: Interactive 3D Gaussian Scene Reconstruction via Multi-Scans FusionWenhao Hu, Zesheng Li, Haonan Zhou, Liu Liu 等AAAI 2026
它引用的顶会 Paper19
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 被引用 4,089 次
- Neural Sparse Voxel FieldsLingjie Liu, Jiatao Gu, Kyaw Zaw Lin, Tat-Seng Chua 等NeurIPS 2020 · 被引用 1,535 次
相关 Paper
- UniC-Lift: Unified 3D Instance Segmentation via Contrastive LearningAnkit Dhiman, R. Srinath, Jaswanth Reddy, Lokesh R. Boregowda 等AAAI 2026
- Trace3D: Consistent Segmentation Lifting via Gaussian Instance TracingHongyu Shen, Junfeng Ni, Yixin Chen, Weishuo Li 等ICCV 2025 · 被引用 5 次
- Contrastive Lift: 3D Object Instance Segmentation by Slow-Fast Contrastive FusionYash Bhalgat, Iro Laina, João F. Henriques, Andrea Vedaldi 等NeurIPS 2023 · 被引用 78 次
- SGS-3D: High-Fidelity 3D Instance Segmentation via Reliable Semantic Mask Splitting and GrowingChaolei Wang, Yang Luo, Jing Du, Siyu Chen 等AAAI 2026 · 被引用 1 次
- OmniSeg3D: Omniversal 3D Segmentation via Hierarchical Contrastive LearningHaiyang Ying, Yixuan Yin, Jinzhi Zhang, Fan Wang 等CVPR 2024 · 被引用 32 次
