Human Body Restoration with One-Step Diffusion Model and A New Benchmark
Jue Gong, Jingkai Wang, Zheng Chen, Xin Liu, Hong Gu, Yulun Zhang, Xiaokang Yang
Abstract
Human body restoration, as a specific application of image restoration, is widely applied in practice and plays a vital role across diverse fields. However, thorough research remains difficult, particularly due to the lack of benchmark datasets. In this study, we propose a high-quality dataset automated cropping and filtering (HQ-ACF) pipeline. This pipeline leverages existing object detection datasets and other unlabeled images to automatically crop and filter high-quality human images. Using this pipeline, we constructed a person-based restoration with sophisticated objects and natural activities (PERSONA) dataset, which includes training, validation, and test sets. The dataset significantly surpasses other humanrelated datasets in both quality and content richness. Finally, we propose OSDHuman, a novel one-step diffusion model for human body restoration. Specifically, we propose a high-fidelity image embedder (HFIE) as the prompt generator to better guide the model with low-quality human image information, effectively avoiding misleading prompts. Experimental results show that OSDHuman outperforms existing methods in both visual quality and quantitative metrics. The dataset and code are available at: https: //github.com/gobunu/OSDHuman .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4aadbb05-ee0e-4c71-9aa6-0e379cfbe542Cited by top-tier papers4
- HAODiff: Human-Aware One-Step Diffusion via Dual-Prompt GuidanceJue Gong, Tingyu Yang, Jingkai Wang, Zheng Chen et al.NeurIPS 2025 · 5 citations
- FideDiff: Efficient Diffusion Model for High-Fidelity Image Motion DeblurringXiaoyang Liu, Zhengyan Zhou, Zihang Xu, Jiezhang Cao et al.ICLR 2026 · 5 citations
- Face2Scene: Using Facial Degradation as an Oracle for Diffusion-Based Scene RestorationAmirhossein Kazerouni, Maitreya Suin, Tristan T Aumentado-Armstrong, Sina Honari et al.CVPR 2026
- SODiff:Semantic-Oriented Diffusion Model for JPEG Compression Artifacts RemovalTingyu Yang, Jue Gong, Jinpei Guo, Wenbo Li et al.AAAI 2026
Builds on18
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray et al.ICML 2021 · 6,356 citations
- ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score DistillationZhengyi Wang, Cheng Lu, Yikai Wang, Fan Bao et al.NeurIPS 2023 · 1,498 citations
Related papers
- OSDFace: One-Step Diffusion Model for Face RestorationJingkai Wang, Jue Gong, Lin Zhang, Zheng Chen et al.CVPR 2025
- Visual Persona: Foundation Model for Full-Body Human CustomizationJisu Nam, Soowon Son, Zhan Xu, Jing Shi et al.CVPR 2025
- GeneMAN: Generalizable Single-Image 3D Human Reconstruction from Multi-Source Human DataWentao Wang, Hang Ye, Fangzhou Hong, Xue Yang et al.NeurIPS 2025 · 6 citations
- Person Image Synthesis via Denoising Diffusion ModelAnkan Kumar Bhunia, Salman H. Khan, Hisham Cholakkal, Rao Muhammad Anwer et al.CVPR 2023
- RAP-SR: RestorAtion Prior Enhancement in Diffusion Models for Realistic Image Super-ResolutionJiangang Wang, Qingnan Fan, Jinwei Chen, Hong Gu et al.AAAI 2025 · 4 citations
