InsActor: Instruction-driven Physics-based Characters
Jiawei Ren, Mingyuan Zhang, Cunjun Yu, Xiao Ma, Liang Pan, Ziwei Liu
Abstract
Generating animation of physics-based characters with intuitive control has long been a desirable task with numerous applications. However, generating physically simulated animations that reflect high-level human instructions remains a difficult problem due to the complexity of physical environments and the richness of human language. In this paper, we present InsActor, a principled generative framework that leverages recent advancements in diffusion-based human motion models to produce instruction-driven animations of physics-based characters. Our framework empowers InsActor to capture complex relationships between high-level human instructions and character motions by employing diffusion policies for flexibly conditioned motion planning. To overcome invalid states and infeasible state transitions in planned motions, InsActor discovers low-level skills and maps plans to latent skill sequences in a compact latent space. Extensive experiments demonstrate that InsActor achieves state-of-the-art results on various tasks, including instruction-driven motion generation and instruction-driven waypoint heading. Notably, the ability of InsActor to generate physically simulated animations using high-level human instructions makes it a valuable tool, particularly in executing long-horizon tasks with a rich set of instructions. Our project page is available at jiawei-ren.github.io/projects/insactor * Equal contribution 37th Conference on Neural Information Processing Systems (NeurIPS 2023).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9e2094ca-a87c-471b-9250-8b71b31232afCited by top-tier papers15
- Iterative Motion Editing with Natural LanguagePurvi Goel, Kuan-Chieh Wang, C. Karen Liu, Kayvon FatahalianSIGGRAPH 2024 · 22 citations
- SuperPADL: Scaling Language-Directed Physics-Based Control with Progressive Supervised DistillationJordan Juravsky, Yunrong Guo, Sanja Fidler, Xue Bin PengSIGGRAPH 2024 · 5 citations
- Omni-Supervised Motion Editing: Balancing Change and Invariance through Positive-Negative LearningZhenwu Shi, Jingyu Gong, Peiwei Wang, Xingzan Wang et al.CVPR 2026 · 4 citations
- InterAgent: Physics-based Multi-agent Command Execution via Diffusion on Interaction GraphsBin Li, Ruichi Zhang, Han Liang, Jingyan Zhang et al.CVPR 2026 · 4 citations
- PhysReaction: Physically Plausible Real-Time Humanoid Reaction Synthesis via Forward Dynamics Guided 4D ImitationYunze Liu, Changxi Chen, Chenjing Ding, Li YiACM MM 2024 · 2 citations
Builds on16
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 13,211 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Planning with Diffusion for Flexible Behavior SynthesisMichael Janner, Yilun Du, Joshua B. Tenenbaum, Sergey LevineICML 2022 · 1,115 citations
Related papers
- UniPhys: Unified Planner and Controller with Diffusion for Flexible Physics-Based Character ControlYan Wu, Korrawe Karunratanakul, Zhengyi Luo, Siyu TangICCV 2025 · 3 citations
- PhysAnimator: Physics-Guided Generative Cartoon AnimationTianyi Xie, Yiwei Zhao, Ying Jiang, Chenfanfu JiangCVPR 2025
- Taming Diffusion Probabilistic Models for Character ControlRui Chen, Mingyi Shi, Shaoli Huang, Ping Tan et al.SIGGRAPH 2024 · 30 citations
- IntentMotion: Learning Intent-Aware Human Motion from Language in 3D ScenesWenfeng Song, Shi Zheng, Xinyu Zhang, Xingliang Jin et al.AAAI 2026
- SkillDiffuser: Interpretable Hierarchical Planning via Skill Abstractions in Diffusion-Based Task ExecutionZhixuan Liang, Yao Mu, Hengbo Ma, Masayoshi Tomizuka et al.CVPR 2024
