PerReactor: Offline Personalised Multiple Appropriate Facial Reaction Generation
Hengde Zhu, Xiangyu Kong, Weicheng Xie, Xin Huang, Xilin He, Lu Liu, Linlin Shen, Wei Zhang, Hatice Gunes, Siyang Song
Abstract
In dyadic human-human interactions, individuals may express multiple different facial reactions in response to the same/similar behaviours expressed by their conversational partners depending on their personalised behaviour patterns. As a result, frequently-employed reconstruction loss-based strategies lead the training of previous automatic facial reaction generation (FRG) models to not only suffer from the 'one-to-many mapping' problem, but also fail to comprehensively consider the quality of the generated facial reactions. Besides, none of them considered such personalised behaviour patterns in generating facial reactions. In this paper, we propose the first adversarial FRG model training strategy which jointly learns appropriateness and realism discriminators to provide comprehensive task-specific supervision for training the target facial reaction generators, and reformulates the 'one-to-many (facial reactions) mapping' training problem as a 'one-to-one (distribution) mapping' training task, i.e., the FRG model is trained to output a distribution representing multiple appropriate/plausible facial reaction from each input human behaviour. In addition, our approach also serves as the first offline FRG approach that considers personalised behaviour patterns in generating of target individuals' facial reactions. Experiments show that our PerReactor not only largely outperformed all existing offline solutions for generating appropriate, diverse and realistic facial reactions, but also is the first offline approach that can effectively generate personalised appropriate facial reactions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on5
- Finite Scalar Quantization: VQ-VAE Made SimpleFabian Mentzer, David Minnen, Eirikur Agustsson, Michael TschannenICLR 2024 · 442 citations
- BeLFusion: Latent Diffusion for Behavior-Driven Human Motion PredictionGermán Barquero, Sergio Escalera, Cristina PalmeroICCV 2023 · 107 citations
- Learning to Listen: Modeling Non-Deterministic Dyadic Facial MotionEvonne Ng, Hanbyul Joo, Liwen Hu, Hao Li et al.CVPR 2022 · 87 citations
- Personality Recognition by Modelling Person-specific Cognitive Processes using Graph RepresentationZilong Shao, Siyang Song, Shashank Jaiswal, Linlin Shen et al.ACM MM 2021 · 45 citations
- PerFRDiff: Personalised Weight Editing for Multiple Appropriate Facial Reaction GenerationHengde Zhu, Xiangyu Kong, Weicheng Xie, Xin Huang et al.ACM MM 2024 · 12 citations
Related papers
- Smooth Online Multiple Appropriate Facial Reaction GenerationWeicheng Xie, Chunlin Yan, Siyang Song, Zitong Yu et al.ACM MM 2025
- ReactDiff: Fundamental Multiple Appropriate Facial Reaction Diffusion ModelCheng Luo, Siyang Song, Siyuan Yan, Zhen Yu et al.ACM MM 2025 · 1 citation
- FReeNet: Multi-Identity Face ReenactmentJiangning Zhang, Xianfang Zeng, Mengmeng Wang, Yusu Pan et al.CVPR 2020
- ReGenNet: Towards Human Action-Reaction SynthesisLiang Xu, Yizhou Zhou, Yichao Yan, Xin Jin et al.CVPR 2024 · 18 citations
- Unsupervised Learning Facial Parameter Regressor for Action Unit Intensity Estimation via Differentiable RendererXinhui Song, Tianyang Shi, Zunlei Feng, Mingli Song et al.ACM MM 2020 · 6 citations
