3D Human Texture Estimation from a Single Image with Transformers
Xiangyu Xu, Chen Change Loy
Abstract
We propose a Transformer-based framework for 3D human texture estimation from a single image. The proposed Transformer is able to effectively exploit the global information of the input image, overcoming the limitations of existing methods that are solely based on convolutional neural networks. In addition, we also propose a mask-fusion strategy to combine the advantages of the RGB-based and texture-flow-based models. We further introduce a part-style loss to help reconstruct high-fidelity colors without introducing unpleasant artifacts. Extensive experiments demonstrate the effectiveness of the proposed method against state-of-the-art 3D human texture estimation approaches both quantitatively and qualitatively. The project page is at https://www.mmlab-ntu.com/project/texformer.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ff973621-e80e-4975-bb55-db5a3f6e92e9Cited by top-tier papers8
- DINAR: Diffusion Inpainting of Neural Textures for One-Shot Human AvatarsDavid Svitov, Dmitrii Gudkov, Renat Bashirov, Victor LempitskyICCV 2023 · 37 citations
- Learning Interaction-aware 3D Gaussian Splatting for One-shot Hand AvatarsXuan Huang, Hanhui Li, Wanquan Liu, Xiaodan Liang et al.NeurIPS 2024 · 6 citations
- Human MotionFormer: Transferring Human Motions with Vision TransformersHongyu Liu, Xintong Han, Chenbin Jin, Lihui Qian et al.ICLR 2023 · 5 citations
- ReFu: Refine and Fuse the Unobserved View for Detail-Preserving Single-Image 3D Human ReconstructionGyumin Shim, Minsoo Lee, Jaegul ChooACM MM 2022 · 4 citations
- UVMap-ID: A Controllable and Personalized UV Map Generative ModelWeijie Wang, Jichao Zhang, Chang Liu, Xia Li et al.ACM MM 2024 · 3 citations
Builds on15
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 1,139 citations
- Multi-Garment Net: Learning to Dress 3D People From ImagesBharat Lal Bhatnagar, Garvita Tiwari, Christian Theobalt, Gerard Pons-MollICCV 2019 · 447 citations
- DeepHuman: 3D Human Reconstruction From a Single ImageZerong Zheng, Tao Yu, Yixuan Wei, Qionghai Dai et al.ICCV 2019 · 367 citations
Related papers
- Human Parsing Based Texture Transfer from Single Image to 3D Human via Cross-View ConsistencyFang Zhao, Shengcai Liao, Kaihao Zhang, Ling ShaoNeurIPS 2020 · 22 citations
- 3D Human Mesh Regression With Dense CorrespondenceWang Zeng, Wanli Ouyang, Ping Luo, Wentao Liu et al.CVPR 2020
- Sampling is Matter: Point-Guided 3D Human Mesh ReconstructionJeonghwan Kim, Mi-Gyeong Gwon, Hyunwoo Park, Hyukmin Kwon et al.CVPR 2023
- Deformable Mesh Transformer for 3D Human Mesh RecoveryYusuke YoshiyasuCVPR 2023
- 3D Human Pose Estimation with Spatial and Temporal TransformersCe Zheng, Sijie Zhu, Matías Mendieta, Taojiannan Yang et al.ICCV 2021 · 648 citations
