Knowledge-inspired 3D Scene Graph Prediction in Point Cloud
Shoulong Zhang, Shuai Li, Aimin Hao, Hong Qin
Abstract
Prior commonsense knowledge integration helps identify semantic entities and their relationships in a graphical representation, however, its meaningful abstraction and intervention remain elusive. This paper advocates a knowledge-inspired 3D scene graph prediction method solely based on point clouds. At the mathematical modeling level, we formulate the task as two sub-problems: commonsense learning and scene graph prediction with learned prior knowledge. Unlike conventional methods that learn knowledge embedding and regular patterns from encoded visual information, we propose to suppress the misunderstandings caused by appearance similarities and other perceptual confusion. At the network design level, we devise a graph auto-encoder to automatically extract class-dependent representations and topological patterns from the one-hot class labels and their intrinsic graphical structures, so that the prior knowledge can avoid perceptual errors and noises. We further devise a scene graph prediction model to predict credible relationship triplets by incorporating the related prototype knowledge with perceptual information. Comprehensive experiments confirm that, our method can successfully learn representative knowledge embedding, and the obtained prior knowledge can effectively enhance the accuracy of relationship predictions. Our thorough evaluations indicate the new method can achieve the state-of-the-art performance compared with other scene graph prediction methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fdf0efd9-ee5b-4512-95df-e981cc361434Cited by top-tier papers16
- 4D Panoptic Scene Graph GenerationJingkang Yang, Jun Cen, Wenxuan Peng, Shuai Liu et al.NeurIPS 2023 · 33 citations
- SGAligner: 3D Scene Alignment with Scene GraphsSayan Deb Sarkar, Ondrej Miksik, Marc Pollefeys, Daniel Barath et al.ICCV 2023 · 27 citations
- MomaGraph: State-Aware Unified Scene Graphs with Vision-Language Models for Embodied Task PlanningYuanchen Ju, Yongyuan Liang, Yen-Jen Wang, Nandiraju Gireesh et al.ICLR 2026 · 5 citations
- SG-PGM: Partial Graph Matching Network with Semantic Geometric Fusion for 3D Scene Graph Alignment and its Downstream TasksYaxu Xie, Alain Pagani, Didier StrickerCVPR 2024 · 5 citations
- Object-Centric Representation Learning for Enhanced 3D Semantic Scene Graph PredictionKunHo Heo, Gihyun Kim, SuYeon Kim, MyeongAh ChoNeurIPS 2025 · 4 citations
Builds on5
- 3D Scene Graph: A Structure for Unified Semantics, 3D Space, and CameraIro Armeni, Zhi-Yang He, Amir Zamir, JunYoung Gwak et al.ICCV 2019 · 474 citations
- RIO: 3D Object Instance Re-Localization in Changing Indoor EnvironmentsJohanna Wald, Armen Avetisyan, Nassir Navab, Federico Tombari et al.ICCV 2019 · 233 citations
- Classification by Attention: Scene Graph Classification with Prior KnowledgeSahand Sharifzadeh, Sina Moayed Baharlou, Volker TrespAAAI 2021 · 61 citations
- Learning 3D Semantic Scene Graphs From 3D Indoor ReconstructionsJohanna Wald, Helisa Dhamo, Nassir Navab, Federico TombariCVPR 2020
- PF-Net: Point Fractal Network for 3D Point Cloud CompletionZitian Huang, Yikuan Yu, Jiawen Xu, Feng Ni et al.CVPR 2020
Related papers
- Multi-view Invariance Learning for 3D Scene Graph Pre-training via Collaborative Cross-Modal RegularizationYucheng Huang, Luping Ji, Ruijie Xiao, Jiayuan SunAAAI 2026
- One-shot Scene Graph GenerationYuyu Guo, Jingkuan Song, Lianli Gao, Heng Tao ShenACM MM 2020 · 26 citations
- CommonScenes: Generating Commonsense 3D Indoor Scenes with Scene GraphsGuangyao Zhai, Evin Pinar Örnek, Shun-Cheng Wu, Yan Di et al.NeurIPS 2023 · 76 citations
- Progressive Seed Generation Auto-encoder for Unsupervised Point Cloud LearningJuYoung Yang, Pyunghwan Ahn, Doyeon Kim, Haeil Lee et al.ICCV 2021 · 24 citations
- CAKE: A Scalable Commonsense-Aware Framework For Multi-View Knowledge Graph CompletionGuanglin Niu, Bo Li, Yongfei Zhang, Shiliang PuACL 2022 · 56 citations
