Duolando: Follower GPT with Off-Policy Reinforcement Learning for Dance Accompaniment
Li Siyao, Tianpei Gu, Zhitao Yang, Zhengyu Lin, Ziwei Liu, Henghui Ding, Lei Yang, Chen Change Loy
摘要
We introduce a novel task within the field of 3D dance generation, termed dance accompaniment, which necessitates the generation of responsive movements from a dance partner, the"follower", synchronized with the lead dancer's movements and the underlying musical rhythm. Unlike existing solo or group dance generation tasks, a duet dance scenario entails a heightened degree of interaction between the two participants, requiring delicate coordination in both pose and position. To support this task, we first build a large-scale and diverse duet interactive dance dataset, DD100, by recording about 117 minutes of professional dancers' performances. To address the challenges inherent in this task, we propose a GPT-based model, Duolando, which autoregressively predicts the subsequent tokenized motion conditioned on the coordinated information of the music, the leader's and the follower's movements. To further enhance the GPT's capabilities of generating stable results on unseen conditions (music and leader motions), we devise an off-policy reinforcement learning strategy that allows the model to explore viable trajectories from out-of-distribution samplings, guided by human-defined rewards. Based on the collected dataset and proposed method, we establish a benchmark with several carefully designed metrics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- MEGADance: Mixture-of-Experts Architecture for Genre-Aware 3D Dance GenerationKaixing Yang, Xulong Tang, Ziqiao Peng, Yuxuan Hu 等NeurIPS 2025 · 被引用 25 次
- 🎧MOSPA: Human Motion Generation Driven by Spatial AudioShuyang Xu, Zhiyang Dou, Mingyi Shi, Liang Pan 等NeurIPS 2025 · 被引用 13 次
- ChoreoCraft: In-situ Crafting of Choreography in Virtual Reality through Creativity Support ToolHyunyoung Han, Kyungeun Jung, Sang Ho YoonCHI 2025 · 被引用 12 次
- DuetGen: Music Driven Two-Person Dance Generation via Hierarchical Masked ModelingAnindita Ghosh, Bing Zhou, Rishabh Dabral, Jian Wang 等SIGGRAPH 2025 · 被引用 11 次
- DyaDiT: A Multi-Modal Diffusion Transformer for Socially Favorable Dyadic Gesture GenerationYICHEN PENG, Jyun-Ting Song, Siyeol Jung, RUOFAN LIU 等CVPR 2026 · 被引用 7 次
它引用的顶会 Paper24
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- AMASS: Archive of Motion Capture As Surface ShapesNaureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll 等ICCV 2019 · 被引用 1,784 次
- AI Choreographer: Music Conditioned 3D Dance Generation with AIST++Ruilong Li, Shan Yang, David A. Ross, Angjoo KanazawaICCV 2021 · 被引用 701 次
- MotionGPT: Human Motion as a Foreign LanguageBiao Jiang, Xin Chen, Wen Liu, Jingyi Yu 等NeurIPS 2023 · 被引用 698 次
- Action-Conditioned 3D Human Motion Synthesis with Transformer VAEMathis Petrovich, Michael J. Black, Gül VarolICCV 2021 · 被引用 672 次
相关 Paper
- MDD: A Dataset for Text-and-Music Conditioned Duet Dance GenerationPrerit Gupta, Jason Alexander Fotso-Puepi, Zhengyuan Li, Jay Mehta 等ICCV 2025 · 被引用 1 次
- Bailando: 3D Dance Generation by Actor-Critic GPT with Choreographic MemoryLi Siyao, Weijiang Yu, Tianpei Gu, Chunze Lin 等CVPR 2022 · 被引用 170 次
- Music-Driven Group ChoreographyNhat Le, Trong-Thang Pham, Tuong Do, Erman Tjiputra 等CVPR 2023
- Dance with You: The Diversity Controllable Dancer Generation via Diffusion ModelsSiyue Yao, Mingjie Sun, Bingliang Li, Fengyu Yang 等ACM MM 2023 · 被引用 23 次
- Explore 3D Dance Generation via Reward Model from Automatically-Ranked DemonstrationsZilin Wang, Haolin Zhuang, Lu Li, Yinmin Zhang 等AAAI 2024 · 被引用 5 次
