IDOL: Instant Photorealistic 3D Human Creation from a Single Image
Yiyu Zhuang, Jiaxi Lv, Hao Wen, Qing Shuai, Ailing Zeng, Hao Zhu, Shifeng Chen, Yujiu Yang, Xun Cao, Wei Liu
Abstract
Figure 1. This work introduces (a) IDOL, a feed-forward, single-image human reconstruction framework that is fast, high-fidelity, and generalizable; (b) Utilizing the proposed Large Generated Human Multi-View Dataset consisting of 100K multi-view subjects, our method exhibits exceptional generalizability in handling diverse human shapes, cross-domain data, severe viewpoints, and occlusions; (c) With a uniform structured representation, the avatars can be directly animatable and easily editable.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2c426261-e3c4-494e-903c-37d327721edaCited by top-tier papers17
- GUAVA: Generalizable Upper Body 3D Gaussian AvatarDongbin Zhang, Yunfei Liu, Lijian Lin, Ye Zhu et al.ICCV 2025 · 9 citations
- PERSONA: Personalized Whole-Body 3D Avatar with Pose-Driven Deformations from a Single ImageGeonhee Sim, Gyeongsik MoonICCV 2025 · 6 citations
- ETCH: Generalizing Body Fitting to Clothed Humans Via Equivariant TightnessBoqian Li, Haiwen Feng, Zeyu Cai, Michael J. Black et al.ICCV 2025 · 5 citations
- Large-scale Codec Avatars: The Unreasonable Effectiveness of Large-scale Avatar PretrainingJunxuan Li, Rawal Khirodkar, Egor Zakharov, Jihyun Lee et al.CVPR 2026 · 3 citations
- Bringing Your Portrait to 3D PresenceJiawei Zhang, Lei Chu, Jiahao Li, Zhenyu Zang et al.CVPR 2026 · 3 citations
Builds on43
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann et al.ICLR 2024 · 4,569 citations
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- LRM: Large Reconstruction Model for Single Image to 3DYicong Hong, Kai Zhang, Jiuxiang Gu, Sai Bi et al.ICLR 2024 · 813 citations
Related papers
- HumanNOVA: Photorealistic, Universal and Rapid 3D Human Avatar Modeling from a Single ImageHezhen Hu, Wangbo Zhao, Lanqing Guo, Hanwen Jiang et al.CVPR 2026
- High-Quality Full-Head 3D Avatar Generation from Any Single Portrait ImageYujie Gao, Chencheng Wang, Xianbing Sun, Jiahui Zhan et al.AAAI 2026
- HumanRAM: Feed-forward Human Reconstruction and Animation Model using TransformersZhiyuan Yu, Zhe Li, Hujun Bao, Can Yang et al.SIGGRAPH 2025 · 2 citations
- GeneMAN: Generalizable Single-Image 3D Human Reconstruction from Multi-Source Human DataWentao Wang, Hang Ye, Fangzhou Hong, Xue Yang et al.NeurIPS 2025 · 6 citations
- LIFe-GoM: Generalizable Human Rendering with Learned Iterative Feedback Over Multi-Resolution Gaussians-on-MeshJing Wen, Alexander G. Schwing, Shenlong WangICLR 2025
