Smooth Online Multiple Appropriate Facial Reaction Generation
Weicheng Xie, Chunlin Yan, Siyang Song, Zitong Yu, Linlin Shen, Laizhong Cui
摘要
In dyadic interactions, facial reactions are crucial for conveying an individuals' responses to their conversational partners. Individuals may exhibit varied but appropriate facial reactions (AFRs) when perceiving the same behavioral expression. Although some recent methods can already respond multiple appropriate facial reactions to the given human speaker behaviors, the AFRs generated by these methods often fail to adequately preserve crucial head motions, leading to visual jitter and unnatural transitions between generated AFR segments. In this paper, we propose a novel and generic PFLPosNet framework which addresses the aforementioned problems at both pre-processing and post-processing stages, where a new pose-aware face behavior localization method PFL is introduced to retain the head pose displacement information from the source data. In addition, the framework proposes a real-time head pose adjustment method, PosNet, to ensure continuity and smoothness in the visual output of the model when using data with correct head pose displacement. Experimental results demonstrate that our approach not only generates more coherent and natural facial reaction sequences but also significantly outperforms existing online MAFRG methods in terms of continuity and smoothness. Our code is made available at https://github.com/rainforcetime/PFLPosNet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper14
- AD-NeRF: Audio Driven Neural Radiance Fields for Talking Head SynthesisYudong Guo, Keyu Chen, Sen Liang, Yong-Jin Liu 等ICCV 2021 · 被引用 510 次
- PIRenderer: Controllable Portrait Image Generation via Semantic Neural RenderingYurui Ren, Ge Li, Yuanqi Chen, Thomas H. Li 等ICCV 2021 · 被引用 284 次
- One-Shot Talking Face Generation from Single-Speaker Audio-Visual Correlation LearningSuzhen Wang, Lincheng Li, Yu Ding, Xin YuAAAI 2022 · 被引用 142 次
- BeLFusion: Latent Diffusion for Behavior-Driven Human Motion PredictionGermán Barquero, Sergio Escalera, Cristina PalmeroICCV 2023 · 被引用 107 次
- FaceVerse: a Fine-grained and Detail-controllable 3D Face Morphable Model from a Hybrid DatasetLizhen Wang, Zhiyuan Chen, Tao Yu, Chenguang Ma 等CVPR 2022 · 被引用 87 次
相关 Paper
- PerFRDiff: Personalised Weight Editing for Multiple Appropriate Facial Reaction GenerationHengde Zhu, Xiangyu Kong, Weicheng Xie, Xin Huang 等ACM MM 2024 · 被引用 12 次
- ReactDiff: Fundamental Multiple Appropriate Facial Reaction Diffusion ModelCheng Luo, Siyang Song, Siyuan Yan, Zhen Yu 等ACM MM 2025 · 被引用 1 次
- PerReactor: Offline Personalised Multiple Appropriate Facial Reaction GenerationHengde Zhu, Xiangyu Kong, Weicheng Xie, Xin Huang 等AAAI 2025 · 被引用 3 次
- PolySLGen: Online Multimodal Speaking-Listening Reaction Generation in Polyadic InteractionZhi-Yi Lin, Thomas Markhorst, Jouh Yeong Chew, Xucong ZhangCVPR 2026 · 被引用 3 次
- SadTalker: Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face AnimationWenxuan Zhang, Xiaodong Cun, Xuan Wang, Yong Zhang 等CVPR 2023
