Let's Chorus: Partner-aware Hybrid Song-Driven 3D Head Animation
Xiumei Xie, Zikai Huang, Wenhao Xu, Peng Xiao, Xuemiao Xu, Huaidong Zhang
2025年份
摘要
Figure 1. Multi-singers Animation. (a): Previous methods construct the 3D facial animation conditioned with an input of single-person audio. (b): With a hybrid song from multi-singers, we argue that it is essential to construct the emotional interaction between each singer for accurate 3D head generation. Motivated by this, we propose the PaChorus framework conditioned on a segment of mixed audio consisting of background music and vocals from multi-singers. With inter-singer interaction modeling, our method can generate emotion-consistent animation sequences.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper16
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech RepresentationsAlexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, Michael AuliNeurIPS 2020 · 被引用 9,451 次
- VAD: Vectorized Scene Representation for Efficient Autonomous DrivingBo Jiang, Shaoyu Chen, Qing Xu, Bencheng Liao 等ICCV 2023 · 被引用 602 次
- AD-NeRF: Audio Driven Neural Radiance Fields for Talking Head SynthesisYudong Guo, Keyu Chen, Sen Liang, Yong-Jin Liu 等ICCV 2021 · 被引用 510 次
- FaceFormer: Speech-Driven 3D Facial Animation with TransformersYingruo Fan, Zhaojiang Lin, Jun Saito, Wenping Wang 等CVPR 2022 · 被引用 218 次
- EmoTalk: Speech-Driven Emotional Disentanglement for 3D Face AnimationZiqiao Peng, Haoyu Wu, Zhenbo Song, Hao Xu 等ICCV 2023 · 被引用 192 次
相关 Paper
- Emotional Voice PuppetryYe Pan, Ruisi Zhang, Shengran Cheng, Shuai Tan 等IEEE VR 2023 · 被引用 21 次
- Talking Together: Synthesizing Co-Located 3D Conversations from AudioMengyi Shan, Shouchieh Chang, Ziqian Bai, Shichen Liu 等CVPR 2026
- KeyFace: Expressive Audio-Driven Facial Animation for Long Sequences via KeyFrame InterpolationAntoni Bigata Casademunt, Michal Stypulkowski, Rodrigo Mira, Stella Bounareli 等CVPR 2025
- MEDTalk: Multimodal Controlled 3D Facial Animation with Dynamic Emotions by Disentangled EmbeddingChang Liu, Ye Pan, Chenyang Ding, Susanto Rahardja 等ACM MM 2025 · 被引用 3 次
- VASA-Rig: Audio-Driven 3D Facial Animation with 'Live' Mood Dynamics in Virtual RealityYe Pan, Chang Liu, Sicheng Xu, Shuai Tan 等IEEE VR 2025 · 被引用 5 次
