OHTA: One-shot Hand Avatar via Data-driven Implicit Priors
Xiaozheng Zheng, Chao Wen, Zhuo Su, Zeran Xu, Zhaohu Li, Yang Zhao, Zhou Xue
Abstract
In this paper, we delve into the creation of one-shot hand avatars, attaining high-fidelity and drivable hand represen-tations swiftly from a single image. With the burgeoning domains of the digital human, the need for quick and per-sonalized hand avatar creation has become increasingly critical. Existing techniques typically require extensive in-put data and may prove cumbersome or even impractical in certain scenarios. To enhance accessibility, we present a novel method OHTA (One-shot Hand avaTAr) that en-ables the creation of detailed hand avatars from merely one image. OHTA tackles the inherent difficulties of this data-limited problem by learning and utilizing data-driven hand priors. Specifically, we design a hand prior model initially employed for 1) learning various hand priors with available data and subsequently for 2) the inversion and fitting of the target identity with prior knowledge. OHTA demonstrates the capability to create high-fidelity hand avatars with con-sistent animatable quality, solely relying on a single image. Furthermore, we illustrate the versatility of OHTA through diverse applications, encompassing text-to-avatar conver-sion, hand editing, and identity latent space manipulation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a8d53f74-4548-4425-8afa-17f67436e264Cited by top-tier papers7
- HumanSplat: Generalizable Single-Image Human Gaussian Splatting with Structure PriorsPanwang Pan, Zhuo Su, Chenguo Lin, Zhen Fan et al.NeurIPS 2024 · 76 citations
- Learning Interaction-aware 3D Gaussian Splatting for One-shot Hand AvatarsXuan Huang, Hanhui Li, Wanquan Liu, Xiaodan Liang et al.NeurIPS 2024 · 6 citations
- Glove2Hand: Synthesizing Natural Hand-Object Interaction from Multi-Modal Sensing GlovesXinyu Zhang, Ziyi Kou, Chuan Qin, Mia Huang et al.CVPR 2026 · 5 citations
- FlexAvatar: Flexible Large Reconstruction Model for Animatable Gaussian Head Avatars with Detailed DeformationCheng Peng, Zhuo Su, Liao Wang, Chen Guo et al.CVPR 2026 · 2 citations
- SRHand: Super-Resolving Hand Images and 3D Shapes via View/Pose-aware Neural Image Representations and Explicit MeshesMinje Kim, Tae-Kyun KimNeurIPS 2025 · 2 citations
Builds on54
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen et al.NeurIPS 2021 · 2,126 citations
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov et al.ICCV 2023 · 1,662 citations
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
Related papers
- GASP: Gaussian Avatars with Synthetic PriorsJack R. Saunders, Charlie Hewitt, Yanan Jian, Marek Kowalski et al.CVPR 2025
- Authentic volumetric avatars from a phone scanChen Cao, Tomas Simon, Jin Kyu Kim, Gabe Schwartz et al.SIGGRAPH 2022 · 123 citations
- FLASHand: Feed-forward reLightable and Animatable Single-view Hand ReconstructionLing-Xiao Zhang, Lin Gao, Wei-Hong He, Yu-Xuan Yang et al.SIGGRAPH 2026
- HumanNOVA: Photorealistic, Universal and Rapid 3D Human Avatar Modeling from a Single ImageHezhen Hu, Wangbo Zhao, Lanqing Guo, Hanwen Jiang et al.CVPR 2026
- HARP: Personalized Hand Reconstruction from a Monocular RGB VideoKorrawe Karunratanakul, Sergey Prokudin, Otmar Hilliges, Siyu TangCVPR 2023
