FactorMatte: Redefining Video Matting for Re-Composition Tasks
Zeqi Gu, Wenqi Xian, Noah Snavely, Abe Davis
Abstract
We propose Factor Matting , an alternative formulation of the video matting problem in terms of counterfactual video synthesis that is better suited for re-composition tasks. The goal of factor matting is to separate the contents of a video into independent components, each representing a counterfactual version of the scene where the contents of other components have been removed. We show that factor matting maps well to a more general Bayesian framing of the matting problem that accounts for complex conditional interactions between layers. Based on this observation, we present a method for solving the factor matting problem that learns augmented patch-based appearance priors to produce useful decompositions even for video with complex cross-layer interactions like splashes, shadows, and reflections. Our method is trained per-video and does not require external training data or any knowledge about the 3D structure of the scene. Through extensive experiments, we show that it is able to produce useful decompositions of scenes with such complex interactions while performing competitively on classical matting tasks as well. We also demonstrate the benefits of our approach on a wide range of downstream video editing tasks. Our project website is at: https://factormatte.github.io/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- BiMatting: Efficient Video Matting via BinarizationHaotong Qin, Lei Ke, Xudong Ma, Martin Danelljan et al.NeurIPS 2023 · 28 citations
- DiffDecompose: Layer-Wise Decomposition of Alpha-Composited Images via Diffusion TransformersZitong Wang, Hang Zhao, Qianyu Zhou, Xuequan Lu et al.CVPR 2026 · 26 citations
- OmnimatteRF: Robust Omnimatte with 3D Background ModelingGeng Lin, Chen Gao, Jia-Bin Huang, Changil Kim et al.ICCV 2023 · 17 citations
- LayerFlow: A Unified Model for Layer-aware Video GenerationSihui Ji, Hao Luo, Xi Chen, Yuanpeng Tu et al.SIGGRAPH 2025 · 12 citations
- Hashing Neural Video Decomposition with Multiplicative Residuals in Space-TimeCheng-Hung Chan, Cheng-Yang Yuan, Cheng Sun, Hwann-Tzong ChenICCV 2023 · 5 citations
Builds on11
- MODNet: Real-Time Trimap-Free Portrait Matting via Objective DecompositionZhanghan Ke, Jiayu Sun, Kaican Li, Qiong Yan et al.AAAI 2022 · 220 citations
- Indices Matter: Learning to Index for Deep Image MattingHao Lu, Yutong Dai, Chunhua Shen, Songcen XuICCV 2019 · 206 citations
- Natural Image Matting via Guided Contextual AttentionYaoyi Li, Hongtao LuAAAI 2020 · 189 citations
- Context-Aware Image Matting for Simultaneous Foreground and Alpha EstimationQiqi Hou, Feng LiuICCV 2019 · 171 citations
- MarioNette: Self-Supervised Sprite LearningDmitriy Smirnov, Michaël Gharbi, Matthew Fisher, Vitor Guizilini et al.NeurIPS 2021 · 47 citations
Related papers
- Associating Objects and Their Effects in Video through Coordination GamesErika Lu, Forrester Cole, Weidi Xie, Tali Dekel et al.NeurIPS 2022 · 8 citations
- Virtual Multi-Modality Self-Supervised Foreground Matting for Human-Object InteractionBo Xu, Han Huang, Cheng Lu, Ziwen Li et al.ICCV 2021 · 7 citations
- Uncertainty-Guided Face Matting for Occlusion-Aware Face TransformationHyebin Cho, Jaehyup LeeACM MM 2025
- Controllable Attention for Structured Layered Video DecompositionJean-Baptiste Alayrac, João Carreira, Relja Arandjelovic, Andrew ZissermanICCV 2019 · 10 citations
- Generative Omnimatte: Learning to Decompose Video into LayersYao-Chih Lee, Erika Lu, Sarah Rumbley, Michal Geyer et al.CVPR 2025
