Bootstraping Clustering of Gaussians for View-consistent 3D Scene Understanding
Wenbo Zhang, Lu Zhang, Ping Hu, Liqian Ma, Yunzhi Zhuge, Huchuan Lu
Abstract
Injecting semantics into 3D Gaussian Splatting (3DGS) has recently garnered significant attention. While current approaches typically distill 3D semantic features from 2D foundational models (e.g., CLIP and SAM) to facilitate novel view segmentation and semantic understanding, their heavy reliance on 2D supervision can undermine cross-view semantic consistency and necessitate complex data preparation processes, therefore hindering view-consistent scene understanding. In this work, we present FreeGS, an unsupervised semantic-embedded 3DGS framework that achieves view-consistent 3D scene understanding without the need for 2D labels. Instead of directly learning semantic features, we introduce the IDentity-coupled Semantic Field (IDSF) into 3DGS, which captures both semantic representations and view-consistent instance indices for each Gaussian. We optimize IDSF with a two-step alternating strategy: semantics help to extract coherent instances in 3D space, while the resulting instances regularize the injection of stable semantics from 2D space. Additionally, we adopt a 2D-3D joint contrastive loss to enhance the complementarity between view-consistent 3D geometry and rich semantics during the bootstrapping process, enabling FreeGS to uniformly perform tasks such as novel-view semantic segmentation, object selection, and 3D object detection. Extensive experiments on LERF-Mask, 3D-OVS, and ScanNet datasets demonstrate that FreeGS performs comparably to state-of-the-art methods while avoiding the complex data preprocessing workload.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7bc537d4-2c21-4243-8e54-72cbc9800149Cited by top-tier papers5
- SIU3R: Simultaneous Scene Understanding and 3D Reconstruction Beyond Feature AlignmentQi Xu, Dongxu Wei, Lingzhe Zhao, Wenpu Li et al.NeurIPS 2025 · 19 citations
- Segment then Splat: Unified 3D Open-Vocabulary Segmentation via Gaussian SplattingYiren Lu, Yunlai Zhou, Yiran Qiao, Chaoda Song et al.NeurIPS 2025 · 9 citations
- LightSplat: Fast and Memory-Efficient Open-Vocabulary 3D Scene Understanding in Five SecondsJaehun Bang, Jinhyeok Kim, Minji Kim, Seungheon Jeong et al.CVPR 2026 · 5 citations
- DentalGS: Pose-Free 3D Gaussian Splatting from Five Intraoral Images for Novel View SynthesisHonghao Dai, Yuanfeng Zhou, Guangshun Wei, Zhihao Li et al.AAAI 2026
- RPE-PAD: Relative Pose Estimation for Pose-agnostic Anomaly DetectionZhipeng Zhang, Mengzan Qi, Rongkang Ma, Yingying Fang et al.AAAI 2026
Builds on16
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- Language-driven Semantic SegmentationBoyi Li, Kilian Q. Weinberger, Serge J. Belongie, Vladlen Koltun et al.ICLR 2022 · 885 citations
- LERF: Language Embedded Radiance FieldsJustin Kerr, Chung Min Kim, Ken Goldberg, Angjoo Kanazawa et al.ICCV 2023 · 620 citations
Related papers
- ObjectGS: Object-Aware Scene Reconstruction and Scene Understanding via Gaussian SplattingRuijie Zhu, Mulin Yu, Linning Xu, Lihan Jiang et al.ICCV 2025 · 1 citation
- UniC-Lift: Unified 3D Instance Segmentation via Contrastive LearningAnkit Dhiman, R. Srinath, Jaswanth Reddy, Lokesh R. Boregowda et al.AAAI 2026
- PointGS: Semantic-Consistent Unsupervised 3D Point Cloud Segmentation with 3D Gaussian SplattingYixiao Song, Qingyong Li, Wen Wang, Zhicheng YanCVPR 2026 · 4 citations
- FHGS: Feature-Homogenized Gaussian SplattingQigeng Duan, Benyun Zhao, Mingqiao Han, Yijun Huang et al.NeurIPS 2025 · 2 citations
- Pose-Free Omnidirectional Gaussian Splatting for 360-Degree Videos with Consistent Depth PriorsChuanqing Zhuang, Xin Lu, Zehui Deng, Zhengda Lu et al.CVPR 2026
