Monte Carlo Scene Search for 3D Scene Understanding
Shreyas Hampali, Sinisa Stekovic, Sayan Deb Sarkar, Chetan Srinivasa Kumar, Friedrich Fraundorfer, Vincent Lepetit
摘要
Abstract We explore how a general AI algorithm can be used for 3D scene understanding to reduce the need for training data. More exactly, we propose a modification of the Monte Carlo Tree Search (MCTS) algorithm to retrieve objects and room layouts from noisy RGB-D scans. While MCTS was developed as a game-playing algorithm, we show it can also be used for complex perception problems. Our adapted MCTS algorithm has few easy-to-tune hyperparameters and can optimise general losses. We use it to optimise the posterior prob-ability of objects and room layout hypotheses given the RGB-D data. This results in an analysis-by-synthesis approach that explores the solution space by rendering the current solution and comparing it to the RGB-D observations. To perform this exploration even more efficiently, we propose simple changes to the standard MCTS' tree construction and exploration policy. We demonstrate our approach on the ScanNet dataset. Our method often retrieves configurations that are better than some manual annotations, especially on layouts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- MonteFloor: Extending MCTS for Reconstructing Accurate Large-Scale Floor PlansSinisa Stekovic, Mahdi Rad, Friedrich Fraundorfer, Vincent LepetitICCV 2021 · 被引用 43 次
- LiteReality: Graphics-Ready 3D Scene Reconstruction from RGB-D ScansZhening Huang, Xiaoyang Wu, Fangcheng Zhong, Hengshuang Zhao 等NeurIPS 2025 · 被引用 26 次
- Convex Decomposition of Indoor ScenesVaibhav Vavilala, David A. ForsythICCV 2023 · 被引用 11 次
- DDIT: Semantic Scene Completion via Deformable Deep Implicit TemplatesHaoang Li, Jinhu Dong, Binghui Wen, Ming Gao 等ICCV 2023 · 被引用 8 次
- DeepSPF: Spherical SO(3)-Equivariant Patches for Scan-to-CAD EstimationDriton Salihu, Adam Misik, Yuankai Wu, Constantin Patsch 等ICLR 2024 · 被引用 2 次
它引用的顶会 Paper7
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 被引用 1,467 次
- Hypersim: A Photorealistic Synthetic Dataset for Holistic Indoor Scene UnderstandingMike Roberts, Jason Ramapuram, Anurag Ranjan, Atulit Kumar 等ICCV 2021 · 被引用 633 次
- Holistic++ Scene Understanding: Single-View 3D Holistic Scene Parsing and Human Pose Estimation With Human-Object Interaction and Physical CommonsenseYixin Chen, Siyuan Huang, Tao Yuan, Yixin Zhu 等ICCV 2019 · 被引用 130 次
- Joint Embedding of 3D Scan and CAD ObjectsManuel Dahnert, Angela Dai, Leonidas J. Guibas, Matthias NießnerICCV 2019 · 被引用 36 次
- MSeg: A Composite Dataset for Multi-Domain Semantic SegmentationJohn Lambert, Zhuang Liu, Ozan Sener, James Hays 等CVPR 2020
相关 Paper
- Reinforcement Learning and Data-Generation for Syntax-Guided SynthesisJulian Parsert, Elizabeth PolgreenAAAI 2024 · 被引用 7 次
- SeeA*: Efficient Exploration-Enhanced A* Search by Selective SamplingDengwei Zhao, Shikui Tu, Lei XuNeurIPS 2024 · 被引用 4 次
- RandomRooms: Unsupervised Pre-training from Synthetic Shapes and Randomized Layouts for 3D Object DetectionYongming Rao, Benlin Liu, Yi Wei, Jiwen Lu 等ICCV 2021 · 被引用 58 次
- Learning 3D Scene Priors with 2D SupervisionYinyu Nie, Angela Dai, Xiaoguang Han, Matthias NießnerCVPR 2023
- Shape Anchor Guided Holistic Indoor Scene UnderstandingMingyue Dong, Linxi Huan, Hanjiang Xiong, Shuhan Shen 等ICCV 2023 · 被引用 5 次
