FlexiClip: Locality-Preserving Free-Form Character Animation
Anant Khandelwal
Abstract
Animating clipart images with seamless motion while maintaining visual fidelity and temporal coherence presents significant challenges. Existing methods, such as AniClipart, effectively model spatial deformations but often fail to ensure smooth temporal transitions, resulting in artifacts like abrupt motions and geometric distortions. Similarly, text-to-video (T2V) and imageto-video (I2V) models struggle to handle clipart due to the mismatch in statistical properties between natural video and clipart styles. This paper introduces FlexiClip, a novel approach designed to overcome these limitations by addressing the intertwined challenges of temporal consistency and geometric integrity. FlexiClip extends traditional Bézier curve-based trajectory modeling with key innovations: temporal Jacobians to correct motion dynamics incrementally, continuous-time modeling via probability flow ODEs (pfODEs) to mitigate temporal noise, and a flow matching loss inspired by GFlowNet principles to optimize smooth motion transitions. These enhancements ensure coherent animations across complex scenarios involving rapid movements and non-rigid deformations. Extensive experiments validate the effectiveness of FlexiClip in generating animations that are not only smooth and natural but also structurally consistent across diverse clipart types, including humans and animals. By integrating spatial and temporal modeling with pretrained video diffusion models, FlexiClip sets a new standard for high-quality clipart animation, offering robust performance across a wide range of visual content. Project
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9f90f13a-9bce-4e46-97d3-d8308e97fcc4Builds on6
- DreamFusion: Text-to-3D using 2D DiffusionBen Poole, Ajay Jain, Jonathan T. Barron, Ben MildenhallICLR 2023 · 463 citations
- Score-based Generative Modeling through Stochastic Evolution Equations in Hilbert SpacesSungbin Lim, Eun-Bi Yoon, Taehyun Byun, Taewon Kang et al.NeurIPS 2023 · 55 citations
- Inflationary Flows: Calibrated Bayesian Inference with Diffusion-Based ModelsDaniela de Albuquerque, John M. PearsonNeurIPS 2024 · 3 citations
- PyramidFlow: High-Resolution Defect Contrastive Localization Using Pyramid Normalizing FlowJiarui Lei, Xiaobo Hu, Yue Wang, Dong LiuCVPR 2023
- VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion ModelsHaoxin Chen, Yong Zhang, Xiaodong Cun, Menghan Xia et al.CVPR 2024
Related papers
- FlipSketch: Flipping Static Drawings to Text-Guided Sketch AnimationsHmrishav Bandyopadhyay, Yi-Zhe SongCVPR 2025
- FlexTraj: Image-to-Video Generation with Flexible Point Trajectory ControlZhiyuan Zhang, Can Wang, Dongdong Chen, Jing LiaoCVPR 2026 · 7 citations
- NeuroClips: Towards High-fidelity and Smooth fMRI-to-Video ReconstructionZixuan Gong, Guangyin Bao, Qi Zhang, Zhongwei Wan et al.NeurIPS 2024 · 39 citations
- Space-Time Diffusion Features for Zero-Shot Text-Driven Motion TransferDanah Yatim, Rafail Fridman, Omer Bar-Tal, Yoni Kasten et al.CVPR 2024 · 29 citations
- Let Your Image Move with Your Motion! -- Implicit Multi-Object Multi-Motion TransferLi Yuze, Dong Gong, Xiao Cao, Junchao Yuan et al.CVPR 2026 · 3 citations
