Human Body Restoration with One-Step Diffusion Model and A New Benchmark
Jue Gong, Jingkai Wang, Zheng Chen, Xin Liu, Hong Gu, Yulun Zhang, Xiaokang Yang
摘要
Human body restoration, as a specific application of image restoration, is widely applied in practice and plays a vital role across diverse fields. However, thorough research remains difficult, particularly due to the lack of benchmark datasets. In this study, we propose a high-quality dataset automated cropping and filtering (HQ-ACF) pipeline. This pipeline leverages existing object detection datasets and other unlabeled images to automatically crop and filter high-quality human images. Using this pipeline, we constructed a person-based restoration with sophisticated objects and natural activities (PERSONA) dataset, which includes training, validation, and test sets. The dataset significantly surpasses other humanrelated datasets in both quality and content richness. Finally, we propose OSDHuman, a novel one-step diffusion model for human body restoration. Specifically, we propose a high-fidelity image embedder (HFIE) as the prompt generator to better guide the model with low-quality human image information, effectively avoiding misleading prompts. Experimental results show that OSDHuman outperforms existing methods in both visual quality and quantitative metrics. The dataset and code are available at: https: //github.com/gobunu/OSDHuman .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- HAODiff: Human-Aware One-Step Diffusion via Dual-Prompt GuidanceJue Gong, Tingyu Yang, Jingkai Wang, Zheng Chen 等NeurIPS 2025 · 被引用 5 次
- FideDiff: Efficient Diffusion Model for High-Fidelity Image Motion DeblurringXiaoyang Liu, Zhengyan Zhou, Zihang Xu, Jiezhang Cao 等ICLR 2026 · 被引用 5 次
- Face2Scene: Using Facial Degradation as an Oracle for Diffusion-Based Scene RestorationAmirhossein Kazerouni, Maitreya Suin, Tristan T Aumentado-Armstrong, Sina Honari 等CVPR 2026
- SODiff:Semantic-Oriented Diffusion Model for JPEG Compression Artifacts RemovalTingyu Yang, Jue Gong, Jinpei Guo, Wenbo Li 等AAAI 2026
它引用的顶会 Paper18
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray 等ICML 2021 · 被引用 6,356 次
- ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score DistillationZhengyi Wang, Cheng Lu, Yikai Wang, Fan Bao 等NeurIPS 2023 · 被引用 1,498 次
相关 Paper
- OSDFace: One-Step Diffusion Model for Face RestorationJingkai Wang, Jue Gong, Lin Zhang, Zheng Chen 等CVPR 2025
- Visual Persona: Foundation Model for Full-Body Human CustomizationJisu Nam, Soowon Son, Zhan Xu, Jing Shi 等CVPR 2025
- GeneMAN: Generalizable Single-Image 3D Human Reconstruction from Multi-Source Human DataWentao Wang, Hang Ye, Fangzhou Hong, Xue Yang 等NeurIPS 2025 · 被引用 6 次
- Person Image Synthesis via Denoising Diffusion ModelAnkan Kumar Bhunia, Salman H. Khan, Hisham Cholakkal, Rao Muhammad Anwer 等CVPR 2023
- RAP-SR: RestorAtion Prior Enhancement in Diffusion Models for Realistic Image Super-ResolutionJiangang Wang, Qingnan Fan, Jinwei Chen, Hong Gu 等AAAI 2025 · 被引用 4 次
