UMSD: High Realism Motion Style Transfer via Unified Mamba-based Diffusion
Ziyun Qian, Zeyu Xiao, Xingliang Jin, Dingkang Yang, Mingcheng Li, Zhenyi Wu, Dongliang Kou, Peng Zhai, Lihua Zhang
摘要
Motion style transfer is a significant research area in computer vision, enabling the rapid switching of stylistic variations for the same motion in virtual digital humans. This dramatically enhances the richness and realism of motions, making it widely applicable in multimedia contexts such as film, gaming, and the Metaverse. However, most existing methods employ a two-stream structure, which often overlooks the intrinsic relationships between content and style motions, resulting in information loss and misalignment. Additionally, these methods struggle to capture temporal dependencies in long-range motion sequences, resulting in less natural outputs. To address these limitations, we propose a Unified Motion Style Diffusion (UMSD) Framework that simultaneously extracts features from content and style motions, achieving comprehensive information interaction. We also introduce the Motion Style Mamba (MSM) denoiser, which, for the first time in motion style transfer, leverages Mamba's powerful sequence modelling capability to produce more temporally coherent stylized motion sequences. Furthermore, we design a diffusion-based content consistency loss and a style consistency loss to ensure that the framework preserves content motion while effectively learning style motion features. Extensive experiments demonstrate that our approach outperforms State-Of-The-Art (SOTA) methods qualitatively and quantitatively, achieving more realistic and coherent motion style transfer.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- MoST: Motion Style Transformer Between Diverse Action ContentsBoeun Kim, Jungho Kim, Hyung Jin Chang, Jin Young ChoiCVPR 2024
- Style-ERD: Responsive and Coherent Online Motion Style TransferTianxin Tao, Xiaohang Zhan, Zhongquan Chen, Michiel van de PanneCVPR 2022 · 被引用 30 次
- MoCoDiff: A Controllable Autoregressive Diffusion Model for Expressive Motion GenerationWenfeng Song, Xuehan Wang, Shuai Li, Yi Chen 等CVPR 2026
- Realistic Full-Body Motion Generation from Sparse Tracking with State Space ModelKun Dong, Jian Xue, Zehai Niu, Xing Lan 等ACM MM 2024 · 被引用 7 次
- Arbitrary Motion Style Transfer with Multi-Condition Motion Latent Diffusion ModelWenfeng Song, Xingliang Jin, Shuai Li, Chenglizhao Chen 等CVPR 2024 · 被引用 18 次
