Consistent and Controllable Image Animation with Motion Diffusion Models
Xin Ma, Yaohui Wang, Gengyun Jia, Xinyuan Chen, Tien-Tsin Wong, Yuan-Fang Li, Cunjian Chen
2025Year
4Top-tier citations
Abstract
Input image "People running" "Doggy barking" "Flames burning and light snow falling" "Man walking on the road" Animated video frames Temporal axis Figure 1. Image animation from our model. Please visit the project page to visualize the animations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6bf1c4aa-ba34-4b9b-958d-7bacc706ce95Cited by top-tier papers4
- Motionagent: Fine-Grained Controllable Video Generation via Motion Field AgentXinyao Liao, Xianfang Zeng, Liao Wang, Gang Yu et al.ICCV 2025 · 2 citations
- MotiF: Making Text Count in Image Animation with Motion Focal LossShijie Wang, Samaneh Azadi, Rohit Girdhar, Saketh Rambhatla et al.CVPR 2025
- PhysAnimator: Physics-Guided Generative Cartoon AnimationTianyi Xie, Yiwei Zhao, Ying Jiang, Chenfanfu JiangCVPR 2025
- MotionStone: Decoupled Motion Intensity Modulation with Diffusion Transformer for Image-to-Video GenerationShuwei Shi, Biao Gong, Xi Chen, Dandan Zheng et al.CVPR 2025
Builds on37
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
Related papers
- 3D Cinemagraphy from a Single ImageXingyi Li, Zhiguo Cao, Huiqiang Sun, Jianming Zhang et al.CVPR 2023
- Learning Fine-Grained Motion Embedding for Landscape AnimationHongwei Xue, Bei Liu, Huan Yang, Jianlong Fu et al.ACM MM 2021 · 9 citations
- AToM: Aligning Text-to-Motion Model at Event-Level with GPT-4Vision RewardHaonan Han, Xiangzuo Wu, Huan Liao, Zunnan Xu et al.CVPR 2025
- Controllable Animation of Fluid Elements in Still ImagesAniruddha Mahapatra, Kuldeep KulkarniCVPR 2022
- VMC: Video Motion Customization Using Temporal Attention Adaption for Text-to-Video Diffusion ModelsHyeonho Jeong, Geon Yeong Park, Jong Chul YeCVPR 2024
