Video Matting via Consistency-Regularized Graph Neural Networks
Tiantian Wang, Sifei Liu, Yapeng Tian, Kai Li, Ming-Hsuan Yang
Abstract
Learning temporally consistent foreground opacity from videos, i.e., video matting, has drawn great attention due to the blossoming of video conferencing. Previous approaches are built on top of image matting models, which fail in maintaining the temporal coherence when being adapted to videos. They either utilize the optical flow to smooth frame-wise prediction, where the performance is dependent on the selected optical flow model; or naively combine feature maps from multiple frames, which does not model well the correspondence of pixels in adjacent frames. In this paper, we propose to enhance the temporal coherence by Consistency-Regularized Graph Neural Networks (CRGNN) with the aid of a synthesized video matting dataset. CRGNN utilizes Graph Neural Networks (GNN) to relate adjacent frames such that pixels or regions that are incorrectly predicted in one frame can be corrected by leveraging information from its neighboring frames. To generalize our model from synthesized videos to real-world videos, we propose a consistency regularization technique to enforce the consistency on the alpha and foreground when blending them with different backgrounds. To evaluate the efficacy of CRGNN, we further collect a real-world dataset with annotated alpha mattes. Compared with state-of-the-art methods that require hand-crafted trimaps or backgrounds for modeling training, CRGNN generates favorably results with the help of unlabeled real training dataset. The source code and datasets are available at https://github.com/TiantianWang/VideoMatting- CRGNN.git.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e5ed434f-36c1-4d8b-93ce-caae89aabb25Cited by top-tier papers13
- MODNet: Real-Time Trimap-Free Portrait Matting via Objective DecompositionZhanghan Ke, Jiayu Sun, Kaican Li, Qiong Yan et al.AAAI 2022 · 220 citations
- Human Instance Matting via Mutual Guidance and Multi-Instance RefinementYanan Sun, Chi-Keung Tang, Yu-Wing TaiCVPR 2022 · 23 citations
- ALEX: Towards Effective Graph Transfer Learning with Noisy LabelsJingyang Yuan, Xiao Luo, Yifang Qin, Zhengyang Mao et al.ACM MM 2023 · 13 citations
- MatAnyone 2: Scaling Video Matting via a Learned Quality EvaluatorPeiqing Yang, Shangchen Zhou, Kai Hao, Qingyi TaoCVPR 2026 · 7 citations
- Revisiting Context Aggregation for Image MattingQinglin Liu, Xiaoqian Lv, Quanling Meng, Zonglin Li et al.ICML 2024 · 6 citations
Builds on8
- Video Object Segmentation Using Space-Time Memory NetworksSeoung Wug Oh, Joon-Young Lee, Ning Xu, Seon Joo KimICCV 2019 · 845 citations
- Zero-Shot Video Object Segmentation via Attentive Graph Neural NetworksWenguan Wang, Xiankai Lu, Jianbing Shen, David J. Crandall et al.ICCV 2019 · 294 citations
- Boundary-Aware Feature Propagation for Scene SegmentationHenghui Ding, Xudong Jiang, Ai Qun Liu, Nadia Magnenat-Thalmann et al.ICCV 2019 · 283 citations
- Indices Matter: Learning to Index for Deep Image MattingHao Lu, Yutong Dai, Chunhua Shen, Songcen XuICCV 2019 · 206 citations
- Context-Aware Image Matting for Simultaneous Foreground and Alpha EstimationQiqi Hou, Feng LiuICCV 2019 · 171 citations
Related papers
- End-to-end Video Matting with Trimap PropagationWei-Lun Huang, Ming-Sui LeeCVPR 2023
- Deep Video Matting via Spatio-Temporal Alignment and AggregationYanan Sun, Guanzhi Wang, Qiao Gu, Chi-Keung Tang et al.CVPR 2021
- Cascade Image Matting with Deformable Graph RefinementZijian Yu, Xuhui Li, Huijuan Huang, Wen Zheng et al.ICCV 2021 · 16 citations
- Attention-guided Temporally Coherent Video Object MattingYunke Zhang, Chi Wang, Miaomiao Cui, Peiran Ren et al.ACM MM 2021 · 31 citations
- Adaptive Human Matting for Dynamic VideosChung-Ching Lin, Jiang Wang, Kun Luo, Kevin Lin et al.CVPR 2023
