Inertia-Guided Flow Completion and Style Fusion for Video Inpainting
Kaidong Zhang, Jingjing Fu, Dong Liu
Abstract
Physical objects have inertia, which resists changes in the velocity and motion direction. Inspired by this, we introduce inertia prior that optical flow, which reflects object motion in a local temporal window, keeps unchanged in the adjacent preceding or subsequent frame. We propose a flow completion network to align and aggregate flow features from the consecutive flow sequences based on the inertia prior. The corrupted flows are completed under the supervision of customized losses on reconstruction, flow smoothness, and consistent ternary census transform. The completed flows with high fidelity give rise to significant improvement on the video inpainting quality. Nevertheless, the existing flow-guided cross-frame warping methods fail to consider the lightening and sharpness variation across video frames, which leads to spatial incoherence after warping from other frames. To alleviate such problem, we propose the Adaptive Style Fusion Network (ASFN), which utilizes the style information extracted from the valid regions to guide the gradient refinement in the warped regions. Moreover, we design a data simulation pipeline to reduce the training difficulty of ASFN. Extensive experiments show the superiority of our method against the state-of-the-art methods quantitatively and qualitatively. The project page is at https://github.com/hitachinsk/ISVI.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ae1b0ec6-784a-4c3d-88a6-e279b3573414Cited by top-tier papers12
- ProPainter: Improving Propagation and Transformer for Video InpaintingShangchen Zhou, Chongyi Li, Kelvin C. K. Chan, Chen Change LoyICCV 2023 · 205 citations
- GRACE: Loss-Resilient Real-Time Video through Neural CodecsYihua Cheng, Ziyi Zhang, Hanchen Li, Anton Arapin et al.NSDI 2024 · 53 citations
- Mask Propagation for Efficient Video Semantic SegmentationYuetian Weng, Mingfei Han, Haoyu He, Mingjie Li et al.NeurIPS 2023 · 36 citations
- BVINet: Unlocking Blind Video Inpainting With Zero AnnotationsZhiliang Wu, Kerui Chen, Kun Li, Hehe Fan et al.ICCV 2025 · 30 citations
- Semantic-Aware Dynamic Parameter for Video Inpainting TransformerEunhye Lee, Jinsu Yoo, Yunjeong Yang, Sungyong Baik et al.ICCV 2023 · 6 citations
Builds on13
- TSM: Temporal Shift Module for Efficient Video UnderstandingJi Lin, Chuang Gan, Song HanICCV 2019 · 2,049 citations
- Free-Form Video Inpainting With 3D Gated Convolution and Temporal PatchGANYa-Liang Chang, Zhe Yu Liu, Kuan-Ying Lee, Winston H. HsuICCV 2019 · 213 citations
- FuseFormer: Fusing Fine-Grained Information in Transformers for Video InpaintingRui Liu, Hanming Deng, Yangyi Huang, Xiaoyu Shi et al.ICCV 2021 · 165 citations
- Copy-and-Paste Networks for Deep Video InpaintingSungho Lee, Seoung Wug Oh, DaeYeun Won, Seon Joo KimICCV 2019 · 137 citations
- Onion-Peel Networks for Deep Video CompletionSeoung Wug Oh, Sungho Lee, Joon-Young Lee, Seon Joo KimICCV 2019 · 112 citations
Related papers
- Learning to Handle Large Obstructions in Video Frame InterpolationLibo Long, Xiao Hu, Jochen LangACM MM 2024
- Towards An End-to-End Framework for Flow-Guided Video InpaintingZhen Li, Chengze Lu, Jianhua Qin, Chun-Le Guo et al.CVPR 2022 · 136 citations
- Progressive Temporal Feature Alignment Network for Video InpaintingXueyan Zou, Linjie Yang, Ding Liu, Yong Jae LeeCVPR 2021
- Preserving Global and Local Temporal Consistency for Arbitrary Video Style TransferXinxiao Wu, Jialu ChenACM MM 2020 · 14 citations
- DRFusion: Drift-Resilient Temporally Consistent Infrared–Visible Video FusionXingyuan Li, HaoYuan Xu, Shulin Li, Xiang Chen et al.ICML 2026
