MM-3DScene: 3D Scene Understanding by Customizing Masked Modeling with Informative-Preserved Reconstruction and Self-Distilled Consistency
Mingye Xu, Mutian Xu, Tong He, Wanli Ouyang, Yali Wang, Xiaoguang Han, Yu Qiao
摘要
mingyexu.github.io/mm3dscene Figure 1 . How to apply masked modeling for large-scale 3D scenes? (a) Conventional random masked modeling on 3D scenes may cause a high risk of uncertainty.In this figure, a chair and a TV are totally masked, which are extremely difficult to be recovered without any context guidance. (b) Our MM-3DScene exploits local statistics to discover and preserve representative structured points, effectively simplifying the pretext task. At each learning step, our method focuses on restoring regional geometry, and enjoys less ambiguity. Moreover, since unmasked areas are underexplored during reconstruction, the model is encouraged to maintain the intrinsic spatial consistency on unmasked points between different masking ratios, which requires the consistent understanding of unmasked areas.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Segment Any Point Cloud Sequences by Distilling Vision Foundation ModelsYouquan Liu, Lingdong Kong, Jun Cen, Runnan Chen 等NeurIPS 2023 · 被引用 169 次
- A Unified Framework for 3D Scene UnderstandingWei Xu, Chunsheng Shi, Sifan Tu, Xin Zhou 等NeurIPS 2024 · 被引用 25 次
- Fine-grained Image-to-LiDAR Contrastive Distillation with Visual Foundation ModelsYifan Zhang, Junhui HouNeurIPS 2024 · 被引用 9 次
- Multi-View Representation is What You Need for Point-Cloud Pre-TrainingSiming Yan, Chen Song, Youkang Kong, Qixing HuangICLR 2024 · 被引用 6 次
- PointCSP: Cross-Sample Semantic Propagation and Stability Preservation in Self-Supervised Point Cloud LearningXinxing Yu, Ajian Liu, Sunyuan Qiang, Hui Ma 等CVPR 2026 · 被引用 1 次
它引用的顶会 Paper31
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 被引用 3,632 次
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui 等ICCV 2019 · 被引用 3,193 次
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 被引用 2,496 次
相关 Paper
- DiffRF: Rendering-Guided 3D Radiance Field DiffusionNorman Müller, Yawar Siddiqui, Lorenzo Porzi, Samuel Rota Bulò 等CVPR 2023
- 3D Mesh Editing Using Masked LRMsWill Gao, Dilin Wang, Yuchen Fan, Aljaz Bozic 等ICCV 2025 · 被引用 6 次
- Self-Supervised Pre-Training with Masked Shape Prediction for 3D Scene UnderstandingLi Jiang, Zetong Yang, Shaoshuai Shi, Vladislav Golyanik 等CVPR 2023
- Clutter Detection and Removal in 3D Scenes with View-Consistent InpaintingFangyin Wei, Thomas A. Funkhouser, Szymon RusinkiewiczICCV 2023 · 被引用 10 次
- CPCM: Contextual Point Cloud Modeling for Weakly-supervised Point Cloud Semantic SegmentationLizhao Liu, Zhuangwei Zhuang, Shangxin Huang, Xunlong Xiao 等ICCV 2023 · 被引用 31 次
