3D human tongue reconstruction from single "in-the-wild" images
Stylianos Ploumpis, Stylianos Moschoglou, Vasileios Triantafyllou, Stefanos Zafeiriou
Abstract
3D face reconstruction from a single image is a task that has garnered increased interest in the Computer Vision community, especially due to its broad use in a number of applications such as realistic 3D avatar creation, pose invariant face recognition and face hallucination. Since the introduction of the 3D Morphable Model in the late 90's, we witnessed an explosion of research aiming at particularly tackling this task. Nevertheless, despite the increasing level of detail in the 3D face reconstructions from single images mainly attributed to deep learning advances, finer and highly deformable components of the face such as the tongue are still absent from all 3D face models in the literature, although being very important for the realness of the 3D avatar representations. In this work we present the first, to the best of our knowledge, end-to-end trainable pipeline that accurately reconstructs the 3D face together with the tongue. Moreover, we make this pipeline robust in “in-the-wild” images by introducing a novel GAN method tailored for 3D tongue surface generation. Finally, we make publicly available to the community the first diverse tongue dataset, consisting of 1,800 raw scans of 700 individuals varying in gender, age, and ethnicity backgrounds**Project url: www.github.com/steliosploumpis/tongue. As we demonstrate in an extensive series of quantitative as well as qualitative experiments, our model proves to be robust and realistically captures the 3D tongue structure, even in adverse “in-the- wild” conditions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Locally Adaptive Neural 3D Morphable ModelsMichail Tarasiou, Rolandos Alexandros Potamias, Eimear O' Sullivan, Stylianos Ploumpis et al.CVPR 2024 · 2 citations
- FitMe: Deep Photorealistic 3D Morphable Model AvatarsAlexandros Lattas, Stylianos Moschoglou, Stylianos Ploumpis, Baris Gecer et al.CVPR 2023
Builds on6
- PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human DigitizationShunsuke Saito, Zeng Huang, Ryota Natsume, Shigeo Morishima et al.ICCV 2019 · 1,411 citations
- DeepHuman: 3D Human Reconstruction From a Single ImageZerong Zheng, Tao Yu, Yixuan Wei, Qionghai Dai et al.ICCV 2019 · 367 citations
- Neural 3D Morphable Models: Spiral Convolutional Networks for 3D Shape Representation Learning and GenerationGiorgos Bouritsas, Sergiy Bokhnyak, Stylianos Ploumpis, Stefanos Zafeiriou et al.ICCV 2019 · 187 citations
- P-nets: Deep Polynomial Neural NetworksGrigorios G. Chrysos, Stylianos Moschoglou, Giorgos Bouritsas, Yannis Panagakis et al.CVPR 2020
- Unsupervised Learning of Probably Symmetric Deformable 3D Objects From Images in the WildShangzhe Wu, Christian Rupprecht, Andrea VedaldiCVPR 2020
Related papers
- Speech Driven Tongue AnimationSalvador Medina, Denis Tomè, Carsten Stoll, Mark Tiede et al.CVPR 2022 · 14 citations
- Generalizable and Animatable Gaussian Head AvatarXuangeng Chu, Tatsuya HaradaNeurIPS 2024 · 115 citations
- Accurate 3D Face Reconstruction with Facial Component TokensTianke Zhang, Xuangeng Chu, Yunfei Liu, Lijian Lin et al.ICCV 2023 · 38 citations
- GAIA: Generative Animatable Interactive Avatars with Expression-conditioned GaussiansZhengming Yu, Tianye Li, Jingxiang Sun, Omer Shapira et al.SIGGRAPH 2025 · 2 citations
- Single-Shot Implicit Morphable Faces with Consistent Texture ParameterizationConnor Z. Lin, Koki Nagano, Jan Kautz, Eric R. Chan et al.SIGGRAPH 2023 · 14 citations
