UMSD: High Realism Motion Style Transfer via Unified Mamba-based Diffusion
Ziyun Qian, Zeyu Xiao, Xingliang Jin, Dingkang Yang, Mingcheng Li, Zhenyi Wu, Dongliang Kou, Peng Zhai, Lihua Zhang
Abstract
Motion style transfer is a significant research area in computer vision, enabling the rapid switching of stylistic variations for the same motion in virtual digital humans. This dramatically enhances the richness and realism of motions, making it widely applicable in multimedia contexts such as film, gaming, and the Metaverse. However, most existing methods employ a two-stream structure, which often overlooks the intrinsic relationships between content and style motions, resulting in information loss and misalignment. Additionally, these methods struggle to capture temporal dependencies in long-range motion sequences, resulting in less natural outputs. To address these limitations, we propose a Unified Motion Style Diffusion (UMSD) Framework that simultaneously extracts features from content and style motions, achieving comprehensive information interaction. We also introduce the Motion Style Mamba (MSM) denoiser, which, for the first time in motion style transfer, leverages Mamba's powerful sequence modelling capability to produce more temporally coherent stylized motion sequences. Furthermore, we design a diffusion-based content consistency loss and a style consistency loss to ensure that the framework preserves content motion while effectively learning style motion features. Extensive experiments demonstrate that our approach outperforms State-Of-The-Art (SOTA) methods qualitatively and quantitatively, achieving more realistic and coherent motion style transfer.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 07f64e0b-9d91-48ff-bcdf-4f1f2292dcb9Cited by top-tier papers1
Ask how each one uses itRelated papers
- MoST: Motion Style Transformer Between Diverse Action ContentsBoeun Kim, Jungho Kim, Hyung Jin Chang, Jin Young ChoiCVPR 2024
- Style-ERD: Responsive and Coherent Online Motion Style TransferTianxin Tao, Xiaohang Zhan, Zhongquan Chen, Michiel van de PanneCVPR 2022 · 30 citations
- MoCoDiff: A Controllable Autoregressive Diffusion Model for Expressive Motion GenerationWenfeng Song, Xuehan Wang, Shuai Li, Yi Chen et al.CVPR 2026
- Realistic Full-Body Motion Generation from Sparse Tracking with State Space ModelKun Dong, Jian Xue, Zehai Niu, Xing Lan et al.ACM MM 2024 · 7 citations
- Arbitrary Motion Style Transfer with Multi-Condition Motion Latent Diffusion ModelWenfeng Song, Xingliang Jin, Shuai Li, Chenglizhao Chen et al.CVPR 2024 · 18 citations
