Let's Chorus: Partner-aware Hybrid Song-Driven 3D Head Animation
Xiumei Xie, Zikai Huang, Wenhao Xu, Peng Xiao, Xuemiao Xu, Huaidong Zhang
Abstract
Figure 1. Multi-singers Animation. (a): Previous methods construct the 3D facial animation conditioned with an input of single-person audio. (b): With a hybrid song from multi-singers, we argue that it is essential to construct the emotional interaction between each singer for accurate 3D head generation. Motivated by this, we propose the PaChorus framework conditioned on a segment of mixed audio consisting of background music and vocals from multi-singers. With inter-singer interaction modeling, our method can generate emotion-consistent animation sequences.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6c55cc29-839d-4d9f-9327-460131dc0580Builds on16
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech RepresentationsAlexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, Michael AuliNeurIPS 2020 · 9,451 citations
- VAD: Vectorized Scene Representation for Efficient Autonomous DrivingBo Jiang, Shaoyu Chen, Qing Xu, Bencheng Liao et al.ICCV 2023 · 602 citations
- AD-NeRF: Audio Driven Neural Radiance Fields for Talking Head SynthesisYudong Guo, Keyu Chen, Sen Liang, Yong-Jin Liu et al.ICCV 2021 · 510 citations
- FaceFormer: Speech-Driven 3D Facial Animation with TransformersYingruo Fan, Zhaojiang Lin, Jun Saito, Wenping Wang et al.CVPR 2022 · 218 citations
- EmoTalk: Speech-Driven Emotional Disentanglement for 3D Face AnimationZiqiao Peng, Haoyu Wu, Zhenbo Song, Hao Xu et al.ICCV 2023 · 192 citations
Related papers
- Emotional Voice PuppetryYe Pan, Ruisi Zhang, Shengran Cheng, Shuai Tan et al.IEEE VR 2023 · 21 citations
- Talking Together: Synthesizing Co-Located 3D Conversations from AudioMengyi Shan, Shouchieh Chang, Ziqian Bai, Shichen Liu et al.CVPR 2026
- KeyFace: Expressive Audio-Driven Facial Animation for Long Sequences via KeyFrame InterpolationAntoni Bigata Casademunt, Michal Stypulkowski, Rodrigo Mira, Stella Bounareli et al.CVPR 2025
- MEDTalk: Multimodal Controlled 3D Facial Animation with Dynamic Emotions by Disentangled EmbeddingChang Liu, Ye Pan, Chenyang Ding, Susanto Rahardja et al.ACM MM 2025 · 3 citations
- VASA-Rig: Audio-Driven 3D Facial Animation with 'Live' Mood Dynamics in Virtual RealityYe Pan, Chang Liu, Sicheng Xu, Shuai Tan et al.IEEE VR 2025 · 5 citations
