Floating No More: Object-Ground Reconstruction from a Single Image
Yunze Man, Yichen Sheng, Jianming Zhang, Liang-Yan Gui, Yu-Xiong Wang
Abstract
Figure 1. Our proposed ORG (Object Reconstruction with Ground) model simultaneously reconstructs a 3D object, estimates camera parameters, and models the object-ground relationship from a monocular image. During shadow and reflection generation, the prior depthbased object geometry estimation method can result in floating issue or an unnatural shadow on the ground, as demonstrated in red boxes. Our method, on the other hand, achieves significantly more realistic editing and generation, as shown in blue boxes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b0007475-ec98-435b-9f3b-8e473818577dCited by top-tier papers2
- SceneCraft: Layout-Guided 3D Scene GenerationXiuyu Yang, Yunze Man, Jun-Kun Chen, Yu-Xiong WangNeurIPS 2024 · 58 citations
- Shoe Style-Invariant and Ground-Aware Learning for Dense Foot Contact EstimationDaniel Jung, Kyoung Mu LeeCVPR 2026 · 1 citation
Builds on30
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan et al.ICCV 2021 · 4,909 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov et al.ICCV 2023 · 1,662 citations
- ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score DistillationZhengyi Wang, Cheng Lu, Yikai Wang, Fan Bao et al.NeurIPS 2023 · 1,498 citations
Related papers
- VisFusion: Visibility-Aware Online 3D Scene Reconstruction from VideosHuiyu Gao, Wei Mao, Miaomiao LiuCVPR 2023
- MetaShadow: Object-Centered Shadow Detection, Removal, and SynthesisTianyu Wang, Jianming Zhang, Haitian Zheng, Zhihong Ding et al.CVPR 2025
- CUPID: Generative 3D Reconstruction via Joint Object and Pose ModelingBinbin Huang, Haobin Duan, Yiqun Zhao, Zibo Zhao et al.CVPR 2026 · 8 citations
- MoGDE: Boosting Mobile Monocular 3D Object Detection with Ground Depth EstimationYunsong Zhou, Quan Liu, Hongzi Zhu, Yunzhe Li et al.NeurIPS 2022 · 23 citations
- Monocular 3D Object Detection with Decoupled Structured Polygon Estimation and Height-Guided Depth EstimationYingjie Cai, Buyu Li, Zeyu Jiao, Hongsheng Li et al.AAAI 2020 · 100 citations
