Generative Omnimatte: Learning to Decompose Video into Layers
Yao-Chih Lee, Erika Lu, Sarah Rumbley, Michal Geyer, Jia-Bin Huang, Tali Dekel, Forrester Cole
Abstract
Input video Output: Omnimatte layers Object removal Layer editing See-through foreground Motion retiming Layer resizing + background replacement ActionShot (duplicating + retiming) Figure 1 . Generative Omnimatte. Our method decomposes a video into a set of RGBA omnimatte layers, where each layer consists of a fully-visible object and its associated effects like shadows and reflections. We improve upon existing work by adding a generative video prior, allowing our method to complete occluded regions (top, middle) and handle dynamic backgrounds (bottom).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers16
- DiffDecompose: Layer-Wise Decomposition of Alpha-Composited Images via Diffusion TransformersZitong Wang, Hang Zhao, Qianyu Zhou, Xuequan Lu et al.CVPR 2026 · 26 citations
- Generative Video Motion Editing with 3D Point TracksYao-Chih Lee, Zhoutong Zhang, Jiahui Huang, Jui-Hsien Wang et al.CVPR 2026 · 23 citations
- Precise Object and Effect Removal with Adaptive Target-Aware AttentionJixin Zhao, Zhouxia Wang, Peiqing Yang, Shangchen ZhouCVPR 2026 · 13 citations
- LayerFlow: A Unified Model for Layer-aware Video GenerationSihui Ji, Hao Luo, Xi Chen, Yuanpeng Tu et al.SIGGRAPH 2025 · 12 citations
- Object-Centric Latent Action LearningAlbina Klepach, Alexander Nikulin, Ilya Zisman, Denis Tarasov et al.AAAI 2026 · 7 citations
Builds on34
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language ModelsJunnan Li, Dongxu Li, Silvio Savarese, Steven C. H. HoiICML 2023 · 7,873 citations
- SDEdit: Guided Image Synthesis and Editing with Stochastic Differential EquationsChenlin Meng, Yutong He, Yang Song, Jiaming Song et al.ICLR 2022 · 2,128 citations
- MultiDiffusion: Fusing Diffusion Paths for Controlled Image GenerationOmer Bar-Tal, Lior Yariv, Yaron Lipman, Tali DekelICML 2023 · 575 citations
Related papers
- OmnimatteRF: Robust Omnimatte with 3D Background ModelingGeng Lin, Chen Gao, Jia-Bin Huang, Changil Kim et al.ICCV 2023 · 17 citations
- Omnimatte: Associating Objects and Their Effects in VideoErika Lu, Forrester Cole, Tali Dekel, Andrew Zisserman et al.CVPR 2021
- Omnimatte3D: Associating Objects and Their Effects in Unconstrained Monocular VideoMohammed Suhail, Erika Lu, Zhengqi Li, Noah Snavely et al.CVPR 2023
- Generative Video PropagationShaoteng Liu, Tianyu Wang, Jui-Hsien Wang, Qing Liu et al.CVPR 2025
- Generative Image Layer Decomposition with Visual EffectsJinrui Yang, Qing Liu, Yijun Li, Soo Ye Kim et al.CVPR 2025
