Training Matting Models Without Alpha Labels
Wenze Liu, Zixuan Ye, Hao Lu, Zhiguo Cao, Xiangyu Yue
Abstract
The labelling difficulty has been a longstanding problem in deep image matting. To escape from fine labels, this work explores using rough annotations such as trimaps coarsely indicating the foreground/background as supervision. We present that the cooperation between learned semantics from indicated known regions and proper assumed matting rules can help infer alpha values at transition areas. Inspired by the nonlocal principle in traditional image matting, we build a directional distance consistency loss (DDC loss) at each pixel neighborhood to constrain the alpha values conditioned on the input image. DDC loss forces the distance of similar pairs on the alpha matte and on its corresponding image to be consistent. In this way, the alpha values can be propagated from learned known regions to unknown transition areas. With only images and trimaps, a matting model can be trained under the supervision of a known loss and the proposed DDC loss. Experiments on AM-2K and P3M-10K dataset show that our paradigm achieves comparable performance with the fine-label-supervised baseline, while sometimes offers even more satisfying results than human-labelled ground truth. Code is available at https://github.com/ poppuppy/alpha-free-matting .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4fedbf1a-8f5d-4cc3-9a64-92984dc8eeadCited by top-tier papers2
- MatAnyone 2: Scaling Video Matting via a Learned Quality EvaluatorPeiqing Yang, Shangchen Zhou, Kai Hao, Qingyi TaoCVPR 2026 · 7 citations
- MatAnyone: Stable Video Matting with Consistent Memory PropagationPeiqing Yang, Shangchen Zhou, Jixin Zhao, Qingyi Tao et al.CVPR 2025
Builds on14
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- MODNet: Real-Time Trimap-Free Portrait Matting via Objective DecompositionZhanghan Ke, Jiayu Sun, Kaican Li, Qiong Yan et al.AAAI 2022 · 220 citations
- Indices Matter: Learning to Index for Deep Image MattingHao Lu, Yutong Dai, Chunhua Shen, Songcen XuICCV 2019 · 206 citations
- Natural Image Matting via Guided Contextual AttentionYaoyi Li, Hongtao LuAAAI 2020 · 189 citations
Related papers
- Improved Image Matting via Real-Time User Clicks and Uncertainty EstimationTianyi Wei, Dongdong Chen, Wenbo Zhou, Jing Liao et al.CVPR 2021
- DiffusionMat: Alpha Matting as Deterministic Sequential Refinement LearningYangyang Xu, Shengfeng He, Wenqi Shao, Yong Du et al.ACM MM 2025 · 1 citation
- Background Matting: The World Is Your Green ScreenSoumyadip Sengupta, Vivek Jayaram, Brian Curless, Steven M. Seitz et al.CVPR 2020
- Disentangled Image MattingShaofan Cai, Xiaoshuai Zhang, Haoqiang Fan, Haibin Huang et al.ICCV 2019 · 127 citations
- Boosting Semantic Human Matting With Coarse AnnotationsJinlin Liu, Yuan Yao, Wendi Hou, Miaomiao Cui et al.CVPR 2020
