Cross-modal Latent Space Alignment for Image to Avatar Translation
Manuel Ladron de Guevara, Yannick Hold-Geoffroy, Jose Echevarria, Cameron Smith, Yijun Li, Daichi Ito
Abstract
We present a novel method for automatic vectorized avatar generation from a single portrait image. Most existing approaches that create avatars rely on image-to-image translation methods, which present some limitations when applied to 3D rendering, animation, or video. Instead, we leverage modality-specific autoencoders trained on large-scale unpaired portraits and parametric avatars, and then learn a mapping between both modalities via an alignment module trained on a significantly smaller amount of data. The resulting cross-modal latent space preserves facial identity, producing more visually appealing and higher fidelity avatars than previous methods, as supported by our quantitative and qualitative evaluations. Moreover, our method’s virtue of being resolution-independent makes it highly versatile and applicable in a wide range of settings.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on9
- U-GAT-IT: Unsupervised Generative Attentional Networks with Adaptive Layer-Instance Normalization for Image-to-Image TranslationJunho Kim, Minjae Kim, Hyeonwoo Kang, Kwanghee LeeICLR 2020 · 632 citations
- Few-shot Image Generation with Elastic Weight ConsolidationYijun Li, Richard Zhang, Jingwan Lu, Eli ShechtmanNeurIPS 2020 · 193 citations
- Face-to-Parameter Translation for Game Character Auto-CreationTianyang Shi, Yi Yuan, Changjie Fan, Zhengxia Zou et al.ICCV 2019 · 56 citations
- StyleCariGAN: caricature generation via StyleGAN feature map modulationWonjong Jang, Gwangjin Ju, Yucheol Jung, Jiaolong Yang et al.SIGGRAPH 2021 · 51 citations
- Fast and Robust Face-to-Parameter Translation for Game Character Auto-CreationTianyang Shi, Zhengxia Zou, Yi Yuan, Changjie FanAAAI 2020 · 38 citations
Related papers
- Generalizable One-shot 3D Neural Head AvatarXueting Li, Shalini De Mello, Sifei Liu, Koki Nagano et al.NeurIPS 2023 · 12 citations
- PERSE: Personalized 3D Generative Avatars from A Single PortraitHyunsoo Cha, Inhee Lee, Hanbyul JooCVPR 2025
- GAIA: Zero-shot Talking Avatar GenerationTianyu He, Junliang Guo, Runyi Yu, Yuchi Wang et al.ICLR 2024 · 51 citations
- Vid2Avatar-Pro: Authentic Avatar from Videos in the Wild via Universal PriorChen Guo, Junxuan Li, Yash Kant, Yaser Sheikh et al.CVPR 2025
- CADQ: Attribute-Consistent Face Cartoonization with Cross-modal Aligned and Deformable QuantizationYongjie Hu, Yifan Jiang, Ziyun Li, Fei Gao et al.ACM MM 2025
