OccFusion: Rendering Occluded Humans with Generative Diffusion Priors
Adam Sun, Tiange Xiang, Scott L. Delp, Li Fei-Fei, Ehsan Adeli
Abstract
Most existing human rendering methods require every part of the human to be fully visible throughout the input video. However, this assumption does not hold in real-life settings where obstructions are common, resulting in only partial visibility of the human. Considering this, we present OccFusion, an approach that utilizes efficient 3D Gaussian splatting supervised by pretrained 2D diffusion models for efficient and high-fidelity human rendering. We propose a pipeline consisting of three stages. In the Initialization stage, complete human masks are generated from partial visibility masks. In the Optimization stage, human 3D Gaussians are optimized with additional supervision by Score-Distillation Sampling (SDS) to create a complete geometry of the human. Finally, in the Refinement stage, in-context inpainting is designed to further improve rendering quality on the less observed human body parts. We evaluate OccFusion on ZJU-MoCap and challenging OcMotion sequences and find that it achieves state-of-the-art performance in the rendering of occluded humans.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c0a2aea8-854d-4eda-833a-b5d8c4aca82fCited by top-tier papers4
- PhysDiff-VTON: Cross-Domain Physics Modeling and Trajectory Optimization for Virtual Try-OnShibin Mei, Bingbing NiNeurIPS 2025 · 4 citations
- Occlusion-Aware Temporally Consistent Amodal Completion for 3D Human-Object Interaction ReconstructionHyungjun Doh, Dong In Lee, Seunggeun Chi, Pin-Hao Huang et al.ACM MM 2025 · 1 citation
- DeOcc-1-to-3: 3D De-Occlusion from a Single Image via Self-Supervised Multi-View DiffusionYansong Qu, Shaohui Dai, Xinyang Li, Yuze Wang et al.AAAI 2026
- CHROME: Clothed Human Reconstruction with Occlusion-Resilience and Multiview-Consistency from a Single ImageArindam Dutta, Meng Zheng, Zhongpai Gao, Benjamin Planche et al.ICCV 2025
Builds on41
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
Related papers
- CrowdGaussian: Reconstructing High-Fidelity 3D Gaussians for Human Crowd from a Single ImageYizheng Song, Yiyu Zhuang, Qipeng Xu, Haixiang Wang et al.CVPR 2026 · 1 citation
- How to Use Diffusion Priors under Sparse Views?Qisen Wang, Yifan Zhao, Jiawei Ma, Jia LiNeurIPS 2024 · 12 citations
- HumanSplat: Generalizable Single-Image Human Gaussian Splatting with Structure PriorsPanwang Pan, Zhuo Su, Chenguo Lin, Zhen Fan et al.NeurIPS 2024 · 76 citations
- Text-to-3D using Gaussian SplattingZilong Chen, Feng Wang, Yikai Wang, Huaping LiuCVPR 2024
- Disco4D: Disentangled 4D Human Generation and Animation from a Single ImageHui En Pang, Shuai Liu, Zhongang Cai, Lei Yang et al.CVPR 2025
