PerReactor: Offline Personalised Multiple Appropriate Facial Reaction Generation
Hengde Zhu, Xiangyu Kong, Weicheng Xie, Xin Huang, Xilin He, Lu Liu, Linlin Shen, Wei Zhang, Hatice Gunes, Siyang Song
摘要
In dyadic human-human interactions, individuals may express multiple different facial reactions in response to the same/similar behaviours expressed by their conversational partners depending on their personalised behaviour patterns. As a result, frequently-employed reconstruction loss-based strategies lead the training of previous automatic facial reaction generation (FRG) models to not only suffer from the 'one-to-many mapping' problem, but also fail to comprehensively consider the quality of the generated facial reactions. Besides, none of them considered such personalised behaviour patterns in generating facial reactions. In this paper, we propose the first adversarial FRG model training strategy which jointly learns appropriateness and realism discriminators to provide comprehensive task-specific supervision for training the target facial reaction generators, and reformulates the 'one-to-many (facial reactions) mapping' training problem as a 'one-to-one (distribution) mapping' training task, i.e., the FRG model is trained to output a distribution representing multiple appropriate/plausible facial reaction from each input human behaviour. In addition, our approach also serves as the first offline FRG approach that considers personalised behaviour patterns in generating of target individuals' facial reactions. Experiments show that our PerReactor not only largely outperformed all existing offline solutions for generating appropriate, diverse and realistic facial reactions, but also is the first offline approach that can effectively generate personalised appropriate facial reactions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Finite Scalar Quantization: VQ-VAE Made SimpleFabian Mentzer, David Minnen, Eirikur Agustsson, Michael TschannenICLR 2024 · 被引用 442 次
- BeLFusion: Latent Diffusion for Behavior-Driven Human Motion PredictionGermán Barquero, Sergio Escalera, Cristina PalmeroICCV 2023 · 被引用 107 次
- Learning to Listen: Modeling Non-Deterministic Dyadic Facial MotionEvonne Ng, Hanbyul Joo, Liwen Hu, Hao Li 等CVPR 2022 · 被引用 87 次
- Personality Recognition by Modelling Person-specific Cognitive Processes using Graph RepresentationZilong Shao, Siyang Song, Shashank Jaiswal, Linlin Shen 等ACM MM 2021 · 被引用 45 次
- PerFRDiff: Personalised Weight Editing for Multiple Appropriate Facial Reaction GenerationHengde Zhu, Xiangyu Kong, Weicheng Xie, Xin Huang 等ACM MM 2024 · 被引用 12 次
相关 Paper
- Smooth Online Multiple Appropriate Facial Reaction GenerationWeicheng Xie, Chunlin Yan, Siyang Song, Zitong Yu 等ACM MM 2025
- ReactDiff: Fundamental Multiple Appropriate Facial Reaction Diffusion ModelCheng Luo, Siyang Song, Siyuan Yan, Zhen Yu 等ACM MM 2025 · 被引用 1 次
- FReeNet: Multi-Identity Face ReenactmentJiangning Zhang, Xianfang Zeng, Mengmeng Wang, Yusu Pan 等CVPR 2020
- ReGenNet: Towards Human Action-Reaction SynthesisLiang Xu, Yizhou Zhou, Yichao Yan, Xin Jin 等CVPR 2024 · 被引用 18 次
- Unsupervised Learning Facial Parameter Regressor for Action Unit Intensity Estimation via Differentiable RendererXinhui Song, Tianyang Shi, Zunlei Feng, Mingli Song 等ACM MM 2020 · 被引用 6 次
