InterSyn: Interleaved Learning for Dynamic Motion Synthesis in the Wild
Yiyi Ma, Yuanzhi Liang, Xiu Li, Chi Zhang, Xuelong Li
Abstract
We present Interleaved Learning for Motion Synthesis (Inter-Syn), a novel framework that targets the generation of realistic interaction motions by learning from integrated motions that consider both solo and multi-person dynamics. Unlike previous methods that treat these components separately, InterSyn employs an interleaved learning strategy to capture the natural, dynamic interactions and nuanced coordination inherent in real-world scenarios. Our framework comprises two key modules: the Interleaved Interaction Synthesis (INS) module, which jointly models solo and interactive behaviors in a unified paradigm from a first-person perspective to support multiple character interactions, and the Relative Coordination Refinement (REC) module, which refines mutual dynamics and ensures synchronized motions among characters. Experimental results show that the motion sequences generated by InterSyn exhibit higher text-to-motion alignment and improved diversity compared with recent methods, setting a new benchmark for robust and natural motion synthesis. Additionally, our code will be open-sourced in the future to promote further research and development in this area. Project website: https://myy888.github.io/InterSyn/
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5ed4c1e7-a3e6-43b2-a890-588adab904c4Cited by top-tier papers3
- NExT-OMNI: Towards Any-to-Any Omnimodal Foundation Models with Discrete Flow MatchingRun Luo, Xiaobo Xia, Lu Wang, Longze Chen et al.ICLR 2026 · 22 citations
- FrankenMotion: Part-level Human Motion Generation and CompositionChuqiao Li, Xianghui Xie, Yong Cao, Andreas Geiger et al.CVPR 2026 · 10 citations
- Unified Number-Free Text-to-Motion Generation Via Flow MatchingGuanhe Huang, Oya ÇeliktutanCVPR 2026
Builds on14
- Directly Denoising Diffusion ModelsDan Zhang, Jingjing Wang, Feng LuoICML 2024 · 11,724 citations
- AMASS: Archive of Motion Capture As Surface ShapesNaureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll et al.ICCV 2019 · 1,784 citations
- MotionGPT: Human Motion as a Foreign LanguageBiao Jiang, Xin Chen, Wen Liu, Jingyi Yu et al.NeurIPS 2023 · 698 citations
- Generating Diverse and Natural 3D Human Motions from TextChuan Guo, Shihao Zou, Xinxin Zuo, Sen Wang et al.CVPR 2022 · 462 citations
- Action2Motion: Conditioned Generation of 3D Human MotionsChuan Guo, Xinxin Zuo, Sen Wang, Shihao Zou et al.ACM MM 2020 · 394 citations
Related papers
- InterControl: Zero-shot Human Interaction Generation by Controlling Every JointZhenzhi Wang, Jingbo Wang, Yixuan Li, Dahua Lin et al.NeurIPS 2024 · 27 citations
- Text2Interact: High-Fidelity and Diverse Text-to-Two-Person Interaction GenerationQingxuan Wu, Zhiyang Dou, chuan guo, Yiming Huang et al.ICLR 2026 · 10 citations
- Large-Scale Multi-Character Interaction SynthesisZiyi Chang, He Wang, George Alex Koulieris, Hubert P. H. ShumSIGGRAPH 2025 · 2 citations
- Locomotion-Action-Manipulation: Synthesizing Human-Scene Interactions in Complex 3D EnvironmentsJiye Lee, Hanbyul JooICCV 2023 · 55 citations
- Scene-Aware Generative Network for Human Motion SynthesisJingbo Wang, Sijie Yan, Bo Dai, Dahua LinCVPR 2021
