αMatte4K & µMatting: Dataset and Model for Ultra-Micro Precision Alpha Video Matting
Xinyi Chen, Hang Dong, Baowei Jiang, Shenkun Xu, Youqi Guan, Kanle Shi, Kun Gai, Haichuan Song
摘要
High-resolution human video matting aims to predict accurate alpha mattes for semi-transparent regions while ensuring temporal consistency across frames. Despite notable progress, current methods still fail to achieve a satisfactory trade-off between quality and efficiency, with limitations in subject stability, temporal modeling, and computational cost. In this paper, we introduce µMatting, an innovative resolution-agnostic two-stage framework for video matting: (1) coarse matte localization using a portrait-aware masked autoencoder;
(2) refinement of critical regions via sparse 3D convolution, augmented by a temporal modulator that injects global spatio-temporal cues for enhanced consistency and contextual awareness. From data perspective, existing research remains limited by the insufficient quality of datasets, including (1) inaccurate alpha fractional values resulting from imperfect annotation, and (2) visual inconsistencies arising from arbitrary foregroundbackground compositions that lack natural coherence. To address this, we introduce αMatte4K, a large-scale 4Kresolution human video matting dataset, which achieves accurate annotations and physical consistency through physically based rendering (PBR). Extensive experiments show that µMatting surpasses state-of-the-art methods in accuracy and spatio-temporal consistency, while αMatte4K boosts baseline performance, driving applications in realworld scenarios. The project is open-sourced at https: //github.com/kadatec/mu-Matting.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- MODNet: Real-Time Trimap-Free Portrait Matting via Objective DecompositionZhanghan Ke, Jiayu Sun, Kaican Li, Qiong Yan 等AAAI 2022 · 被引用 220 次
- Transparent Image Layer Diffusion using Latent TransparencyLvmin Zhang, Maneesh AgrawalaSIGGRAPH 2024 · 被引用 42 次
- Video Matting via Consistency-Regularized Graph Neural NetworksTiantian Wang, Sifei Liu, Yapeng Tian, Kai Li 等ICCV 2021 · 被引用 31 次
- BiMatting: Efficient Video Matting via BinarizationHaotong Qin, Lei Ke, Xudong Ma, Martin Danelljan 等NeurIPS 2023 · 被引用 28 次
- Real-Time High-Resolution Background MattingShanchuan Lin, Andrey Ryabtsev, Soumyadip Sengupta, Brian L. Curless 等CVPR 2021
相关 Paper
- MaGGIe: Masked Guided Gradual Human Instance MattingChuong Huynh, Seoung Wug Oh, Abhinav Shrivastava, Joon-Young LeeCVPR 2024
- Ultrahigh Resolution Image/Video Matting with Spatio-Temporal SparsityYanan Sun, Chi-Keung Tang, Yu-Wing TaiCVPR 2023
- Generative Video MattingYongtao Ge, Kangyang Xie, Guangkai Xu, Li Ke 等SIGGRAPH 2025 · 被引用 1 次
- Memory Efficient Matting with Adaptive Token RoutingYiheng Lin, Yihan Hu, Chenyi Zhang, Ting Liu 等AAAI 2025 · 被引用 1 次
- Adaptive Human Matting for Dynamic VideosChung-Ching Lin, Jiang Wang, Kun Luo, Kevin Lin 等CVPR 2023
