Adaptive Human Matting for Dynamic Videos
Chung-Ching Lin, Jiang Wang, Kun Luo, Kevin Lin, Linjie Li, Lijuan Wang, Zicheng Liu
Abstract
The most recent efforts in video matting have focused on eliminating trimap dependency since trimap annotations are expensive and trimap-based methods are less adaptable for real-time applications. Despite the latest tripmapfree methods showing promising results, their performance often degrades when dealing with highly diverse and unstructured videos. We address this limitation by introducing Adaptive Matting for Dynamic Videos, termed AdaM, which is a framework designed for simultaneously differentiating foregrounds from backgrounds and capturing alpha matte details of human subjects in the foreground. Two interconnected network designs are employed to achieve this goal: (1) an encoder-decoder network that produces alpha mattes and intermediate masks which are used to guide the transformer in adaptively decoding foregrounds and backgrounds, and (2) a transformer network in which long- and short-term attention combine to retain spatial and temporal contexts, facilitating the decoding of foreground details. We benchmark and study our methods on recently introduced datasets, showing that our model notably improves matting realism and temporal coherence in complex real-world videos and achieves new best-in-class generalizability. Further details and examples are available at https://github.com/microsoft/AdaM.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f51ac00f-25d0-47ae-b074-2648733fa727Cited by top-tier papers6
- MatAnyone 2: Scaling Video Matting via a Learned Quality EvaluatorPeiqing Yang, Shangchen Zhou, Kai Hao, Qingyi TaoCVPR 2026 · 7 citations
- MP-Mat: A 3D-and-Instance-Aware Human Matting and Editing Framework with Multiplane RepresentationSiyi Jiao, Wenzheng Zeng, Yerong Li, Huayu Zhang et al.ICLR 2025
- MatAnyone: Stable Video Matting with Consistent Memory PropagationPeiqing Yang, Shangchen Zhou, Jixin Zhao, Qingyi Tao et al.CVPR 2025
- αMatte4K & µMatting: Dataset and Model for Ultra-Micro Precision Alpha Video MattingXinyi Chen, Hang Dong, Baowei Jiang, Shenkun Xu et al.CVPR 2026
- MaGGIe: Masked Guided Gradual Human Instance MattingChuong Huynh, Seoung Wug Oh, Abhinav Shrivastava, Joon-Young LeeCVPR 2024
Builds on16
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 2,196 citations
- Video Object Segmentation Using Space-Time Memory NetworksSeoung Wug Oh, Joon-Young Lee, Ning Xu, Seon Joo KimICCV 2019 · 845 citations
- Associating Objects with Transformers for Video Object SegmentationZongxin Yang, Yunchao Wei, Yi YangNeurIPS 2021 · 398 citations
- MODNet: Real-Time Trimap-Free Portrait Matting via Objective DecompositionZhanghan Ke, Jiayu Sun, Kaican Li, Qiong Yan et al.AAAI 2022 · 220 citations
- Natural Image Matting via Guided Contextual AttentionYaoyi Li, Hongtao LuAAAI 2020 · 189 citations
Related papers
- Disentangled Image MattingShaofan Cai, Xiaoshuai Zhang, Haoqiang Fan, Haibin Huang et al.ICCV 2019 · 127 citations
- Background Matting: The World Is Your Green ScreenSoumyadip Sengupta, Vivek Jayaram, Brian Curless, Steven M. Seitz et al.CVPR 2020
- OmnimatteRF: Robust Omnimatte with 3D Background ModelingGeng Lin, Chen Gao, Jia-Bin Huang, Changil Kim et al.ICCV 2023 · 17 citations
- Uncertainty-Guided Face Matting for Occlusion-Aware Face TransformationHyebin Cho, Jaehyup LeeACM MM 2025
- Generative Video MattingYongtao Ge, Kangyang Xie, Guangkai Xu, Li Ke et al.SIGGRAPH 2025 · 1 citation
