Temporally Guided Music-to-Body-Movement Generation
Hsuan-Kai Kao, Li Su
摘要
This paper presents a neural network model to generate virtual violinistâĂŹs 3-D skeleton movements from music audio. Improved from the conventional recurrent neural network models for generating 2-D skeleton data in previous works, the proposed model incorporates an encoder-decoder architecture, as well as the selfattention mechanism to model the complicated dynamics in body movement sequences. To facilitate the optimization of self-attention model, beat tracking is applied to determine effective sizes and boundaries of the training examples. The decoder is accompanied with a refining network and a bowing attack inference mechanism to emphasize the right-hand behavior and bowing attack timing. Both objective and subjective evaluations reveal that the proposed model outperforms the state-of-the-art methods. To the best of our knowledge, this work represents the first attempt to generate 3-D violinistsâĂŹ body movements considering key features in musical body movement.
• Computing methodologies → Motion processing; • Applied computing → Media arts; Sound and music computing; • Humancentered computing → Sound-based input / output.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- AI Choreographer: Music Conditioned 3D Dance Generation with AIST++Ruilong Li, Shan Yang, David A. Ross, Angjoo KanazawaICCV 2021 · 被引用 701 次
- Bailando: 3D Dance Generation by Actor-Critic GPT with Choreographic MemoryLi Siyao, Weijiang Yu, Tianpei Gu, Chunze Lin 等CVPR 2022 · 被引用 170 次
- TM2D: Bimodality Driven 3D Dance Generation via Music-Text IntegrationKehong Gong, Dongze Lian, Heng Chang, Chuan Guo 等ICCV 2023 · 被引用 103 次
- Fg-T2M: Fine-Grained Text-Driven Human Motion Generation via Diffusion ModelYin Wang, Zhiying Leng, Frederick W. B. Li, Shun-Cheng Wu 等ICCV 2023 · 被引用 95 次
- Audio Matters Too! Enhancing Markerless Motion Capture with Audio Signals for String Performance CaptureYitong Jin, Zhiping Qiu, Yi Shi, Shuangpeng Sun 等SIGGRAPH 2024 · 被引用 9 次
相关 Paper
- A Human-Computer Duet System for Music PerformanceYuen-Jen Lin, Hsuan-Kai Kao, Yih-Chih Tseng, Ming Tsai 等ACM MM 2020 · 被引用 6 次
- SelfTalk: A Self-Supervised Commutative Training Diagram to Comprehend 3D Talking FacesZiqiao Peng, Yihao Luo, Yue Shi, Hao Xu 等ACM MM 2023 · 被引用 56 次
- EchoAvatar: Real-time Generative Avatar Animation from Audio StreamsBohong Chen, Yumeng Li, Yinglin Xu, Youyi Zheng 等SIGGRAPH 2026
- Audio2Gestures: Generating Diverse Gestures from Speech Audio with Conditional Variational AutoencodersJing Li, Di Kang, Wenjie Pei, Xuefei Zhe 等ICCV 2021 · 被引用 144 次
- Self-supervised Dance Video Synthesis Conditioned on MusicXuanchi Ren, Haoran Li, Zijian Huang, Qifeng ChenACM MM 2020 · 被引用 68 次
