FLNet: Landmark Driven Fetching and Learning Network for Faithful Talking Facial Animation Synthesis
Kuangxiao Gu, Yuqian Zhou, Thomas S. Huang
Abstract
Talking face synthesis has been widely studied in either appearance-based or warping-based methods. Previous works mostly utilize single face image as a source, and generate novel facial animations by merging other person's facial features. However, some facial regions like eyes or teeth, which may be hidden in the source image, can not be synthesized faithfully and stably. In this paper, We present a landmark driven two-stream network to generate faithful talking facial animation, in which more facial details are created, preserved and transferred from multiple source images instead of a single one. Specifically, we propose a network consisting of a learning and fetching stream. The fetching sub-net directly learns to attentively warp and merge facial regions from five source images of distinctive landmarks, while the learning pipeline renders facial organs from the training face space to compensate. Compared to baseline algorithms, extensive experiments demonstrate that the proposed method achieves a higher performance both quantitatively and qualitatively. Codes are at https://github.com/kgu3/FLNet_AAAI2020.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers18
- VirtualCube: An Immersive 3D Video Communication SystemYizhong Zhang, Jiaolong Yang, Zhen Liu, Ruicheng Wang et al.IEEE VR 2022 · 66 citations
- Implicit Warping for Animation with Image SetsArun Mallya, Ting-Chun Wang, Ming-Yu LiuNeurIPS 2022 · 62 citations
- Learned Spatial Representations for Few-shot Talking-Head SynthesisMoustafa Meshry, Saksham Suri, Larry S. Davis, Abhinav ShrivastavaICCV 2021 · 51 citations
- Structure-Aware Motion Transfer with Deformable Anchor ModelJiale Tao, Biao Wang, Borun Xu, Tiezheng Ge et al.CVPR 2022 · 33 citations
- Animating Through Warping: An Efficient Method for High-Quality Facial Expression AnimationZili Yi, Qiang Tang, Vishnu Sanjay Ramiya Srinivasan, Zhan XuACM MM 2020 · 8 citations
Builds on1
Related papers
- That's What I Said: Fully-Controllable Talking Face GenerationYoungjoon Jang, Kyeongha Rho, Jong-Bin Woo, Hyeongkeun Lee et al.ACM MM 2023 · 7 citations
- MODA: Mapping-Once Audio-driven Portrait Animation with Dual AttentionsYunfei Liu, Lijian Lin, Fei Yu, Changyin Zhou et al.ICCV 2023 · 40 citations
- Implicit Identity Representation Conditioned Memory Compensation Network for Talking Head Video GenerationFa-Ting Hong, Dan XuICCV 2023 · 75 citations
- FACIAL: Synthesizing Dynamic Talking Face with Implicit Attribute LearningChenxu Zhang, Yifan Zhao, Yifei Huang, Ming Zeng et al.ICCV 2021 · 149 citations
- Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video GenerationShuling Zhao, Fa-Ting Hong, Xiaoshui Huang, Dan XuCVPR 2025
