UVMap-ID: A Controllable and Personalized UV Map Generative Model
Weijie Wang, Jichao Zhang, Chang Liu, Xia Li, Xingqian Xu, Humphrey Shi, Nicu Sebe, Bruno Lepri
Abstract
Recently, diffusion models have made significant strides in synthesizing realistic 2D human images based on provided text prompts. Building upon this, researchers have extended 2D text-to-image diffusion models into the 3D domain for generating human textures (UV Maps). However, some important problems about UV Map Generative models are still not solved, i.e., how to generate personalized texture maps for any given face image, and how to define and evaluate the quality of these generated texture maps. To solve the above problems, we introduce a novel method, UVMap-ID, which is a controllable and personalized UV Map generative model. Unlike traditional large-scale training methods in 2D, we propose to fine-tune a pre-trained text-to-image diffusion model which is integrated with a face fusion module for achieving ID-driven customized generation. To support the finetuning strategy, we introduce a small-scale attribute-balanced training dataset, including high-quality textures with labeled text and Face ID. Additionally, we introduce some metrics to evaluate the multiple aspects of the textures. Finally, both quantitative and qualitative analyses demonstrate the effectiveness of our method in controllable and personalized UV Map generation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Token Reduction via Local and Global Contexts Optimization for Efficient Video Large Language ModelsJinlong Li, Liyuan Jiang, Haonan Zhang, Nicu SebeCVPR 2026 · 5 citations
- PoInit-of-View: Poisoning Initialization of Views Transfers Across Multiple 3D Reconstruction SystemsWeijie Wang, Songlong Xing, Zhengyu Zhao, Nicu Sebe et al.CVPR 2026 · 1 citation
Builds on33
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Directly Denoising Diffusion ModelsDan Zhang, Jingjing Wang, Feng LuoICML 2024 · 11,724 citations
Related papers
- UV-IDM: Identity-Conditioned Latent Diffusion Model for Face UV-Texture GenerationHong Li, Yutang Feng, Song Xue, Xuhui Liu et al.CVPR 2024
- Texture Generation on 3D Meshes with Point-UV DiffusionXin Yu, Peng Dai, Wenbo Li, Lan Ma et al.ICCV 2023 · 78 citations
- OMGTex: One-stage Multi-style Facial Texture Reconstruction without Geometry GuidanceZitong Xiao, Yuda Qiu, Zisheng Ye, Xiaoguang HanCVPR 2026
- TexGarment: Consistent Garment UV Texture Generation via Efficient 3D Structure-Guided Diffusion TransformerJialun Liu, Jinbo Wu, Xiaobo Gao, Jiakui Hu et al.CVPR 2025
- Paint3D: Paint Anything 3D With Lighting-Less Texture Diffusion ModelsXianfang Zeng, Xin Chen, Zhongqi Qi, Wen Liu et al.CVPR 2024 · 44 citations
