Semi-Supervised Video Inpainting with Cycle Consistency Constraints
Zhiliang Wu, Hanyu Xuan, Changchang Sun, Weili Guan, Kang Zhang, Yan Yan
摘要
Deep learning-based video inpainting has yielded promising results and gained increasing attention from researchers. Generally, these methods assume that the corrupted region masks of each frame are known and easily obtained. However, the annotation of these masks are laborintensive and expensive, which limits the practical application of current methods. Therefore, we expect to relax this assumption by defining a new semi-supervised inpainting setting, making the networks have the ability of completing the corrupted regions of the whole video using the annotated mask of only one frame. Specifically, in this work, we propose an end-to-end trainable framework consisting of completion network and mask prediction network, which are designed to generate corrupted contents of the current frame using the known mask and decide the regions to be filled of the next frame, respectively. Besides, we introduce a cycle consistency loss to regularize the training parameters of these two networks. In this way, the completion network and the mask prediction network can constrain each other, and hence the overall performance of the trained model can be maximized. Furthermore, due to the natural existence of prior knowledge (e.g., corrupted contents and clear borders), current video inpainting datasets are not suitable in the context of semi-supervised video inpainting. Thus, we create a new dataset by simulating the corrupted video of real-world scenarios. Extensive experimental results are reported to demonstrate the superiority of our model in the video inpainting task. Remarkably, although our model is trained in a semi-supervised manner, it can achieve comparable performance as fully-supervised methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- BVINet: Unlocking Blind Video Inpainting With Zero AnnotationsZhiliang Wu, Kerui Chen, Kun Li, Hehe Fan 等ICCV 2025 · 被引用 30 次
- Look Ma, No Hands! Agent-Environment Factorization of Egocentric VideosMatthew Chang, Aditya Prakash, Saurabh GuptaNeurIPS 2023 · 被引用 26 次
- CIRI: Curricular Inactivation for Residue-aware One-shot Video InpaintingWeiying Zheng, Cheng Xu, Xuemiao Xu, Wenxi Liu 等ICCV 2023 · 被引用 12 次
- Elevating Flow-Guided Video Inpainting with Reference GenerationSuhwan Cho, Seoung Wug Oh, Sangyoun Lee, Joon-Young LeeAAAI 2025 · 被引用 2 次
- Blind Bitstream-corrupted Video Recovery via Metadata-guided Diffusion ModelShuyun Wang, Hu Zhang, Xin Shen, Dadong Wang 等CVPR 2025
它引用的顶会 Paper21
- Video Object Segmentation Using Space-Time Memory NetworksSeoung Wug Oh, Joon-Young Lee, Ning Xu, Seon Joo KimICCV 2019 · 被引用 845 次
- RANet: Ranking Attention Network for Fast Video Object SegmentationZiqin Wang, Jun Xu, Li Liu, Fan Zhu 等ICCV 2019 · 被引用 217 次
- Free-Form Video Inpainting With 3D Gated Convolution and Temporal PatchGANYa-Liang Chang, Zhe Yu Liu, Kuan-Ying Lee, Winston H. HsuICCV 2019 · 被引用 213 次
- FuseFormer: Fusing Fine-Grained Information in Transformers for Video InpaintingRui Liu, Hanming Deng, Yangyi Huang, Xiaoyu Shi 等ICCV 2021 · 被引用 165 次
- Copy-and-Paste Networks for Deep Video InpaintingSungho Lee, Seoung Wug Oh, DaeYeun Won, Seon Joo KimICCV 2019 · 被引用 137 次
相关 Paper
- Internal Video Inpainting by Implicit Long-range PropagationHao Ouyang, Tengfei Wang, Qifeng ChenICCV 2021 · 被引用 42 次
- Hierarchical Masked 3D Diffusion Model for Video OutpaintingFanda Fan, Chaoxu Guo, Litong Gong, Biao Wang 等ACM MM 2023 · 被引用 12 次
- An Internal Learning Approach to Video InpaintingHaotian Zhang, Long Mai, Hailin Jin, Zhaowen Wang 等ICCV 2019 · 被引用 77 次
- Single-Stage Semantic Segmentation From Image LabelsNikita Araslanov, Stefan RothCVPR 2020
- Every Frame Counts: Joint Learning of Video Segmentation and Optical FlowMingyu Ding, Zhe Wang, Bolei Zhou, Jianping Shi 等AAAI 2020 · 被引用 80 次
