Guess The Unseen: Dynamic 3D Scene Reconstruction from Partial 2D Glimpses
Inhee Lee, Byungjun Kim, Hanbyul Joo
Abstract
Input: Monocular Video Novel Pose Synthesis Novel View Synthesis Output: Animatable Person 3D-GS Occluded partial observations For Teaser (light) Figure 1. We present a method to reconstruct dynamic scenes from a monocular video capturing partial 2D observations. As a key advantage, our method can estimate the unseen body parts by leveraging a pre-trained diffusion model [40] via SDS method [37]. The reconstructed scenes can be rendered to any viewpoint and each human body can be transformed into any body posture controlled by SMPL [27] parameters.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4272d2fd-b651-457f-9ec2-2272c8a9c484Cited by top-tier papers9
- Walking the Schrödinger Bridge: A Direct Trajectory for Text-to-3D GenerationZiying Li, Xuequan Lu, Xinkui Zhao, Guanjie Cheng et al.NeurIPS 2025 · 4 citations
- HairCUP: Hair Compositional Universal Prior for 3D Gaussian AvatarsByungjun Kim, Shunsuke Saito, Giljoo Nam, Tomas Simon et al.ICCV 2025 · 2 citations
- Occlusion-Aware Temporally Consistent Amodal Completion for 3D Human-Object Interaction ReconstructionHyungjun Doh, Dong In Lee, Seunggeun Chi, Pin-Hao Huang et al.ACM MM 2025 · 1 citation
- Human Interaction-Aware 3D Reconstruction from a Single ImageGwanghyun Kim, Junghun James Kim, Suh Yoon Jeon, Jason Park et al.CVPR 2026
- ShowMak3r: Compositional TV Show ReconstructionSangmin Kim, Seunguk Do, Jaesik ParkCVPR 2025
Builds on39
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
Related papers
- Reconstruct, Inpaint, Test-Time Finetune: Dynamic Novel-view Synthesis from Monocular VideosKaihua Chen, Tarasha Khurana, Deva RamananNeurIPS 2025 · 17 citations
- D^3-Human: Dynamic Disentangled Digital Human from Monocular VideoHonghu Chen, Bo Peng, Yunfan Tao, Juyong ZhangCVPR 2025
- Pseudo-Generalized Dynamic View Synthesis from a VideoXiaoming Zhao, Alex Colburn, Fangchang Ma, Miguel Ángel Bautista et al.ICLR 2024 · 31 citations
- Dynamic View Synthesis as an Inverse ProblemHidir Yesiltepe, Pinar YanardagNeurIPS 2025 · 12 citations
- CHROME: Clothed Human Reconstruction with Occlusion-Resilience and Multiview-Consistency from a Single ImageArindam Dutta, Meng Zheng, Zhongpai Gao, Benjamin Planche et al.ICCV 2025
