Revisiting Context Aggregation for Image Matting
Qinglin Liu, Xiaoqian Lv, Quanling Meng, Zonglin Li, Xiangyuan Lan, Shuo Yang, Shengping Zhang, Liqiang Nie
Abstract
Traditional studies emphasize the significance of context information in improving matting performance. Consequently, deep learning-based matting methods delve into designing pooling or affinity-based context aggregation modules to achieve superior results. However, these modules cannot well handle the context scale shift caused by the difference in image size during training and inference, resulting in matting performance degradation. In this paper, we revisit the context aggregation mechanisms of matting networks and find that a basic encoder-decoder network without any context aggregation modules can actually learn more universal context aggregation, thereby achieving higher matting performance compared to existing methods. Building on this insight, we present AEMatter, a matting network that is straightforward yet very effective. AEMatter adopts a Hybrid-Transformer backbone with appearance-enhanced axis-wise learning (AEAL) blocks to build a basic network with strong context aggregation learning capability. Furthermore, AEMatter leverages a large image training strategy to assist the network in learning context aggregation from data. Extensive experiments on five popular matting datasets demonstrate that the proposed AEMatter outperforms state-of-the-art matting methods by a large margin.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Scaling up Image Segmentation across Data and TasksPei Wang, Zhaowei Cai, Hao Yang, Ashwin Swaminathan et al.CVPR 2025
- Uncertainty-Guided Face Matting for Occlusion-Aware Face TransformationHyebin Cho, Jaehyup LeeACM MM 2025
- Path-Adaptive Matting for Efficient Inference Under Various Computational Cost ConstraintsQinglin Liu, Zonglin Li, Xiaoqian Lv, Xin Sun et al.AAAI 2025
Builds on16
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- On the Variance of the Adaptive Learning Rate and BeyondLiyuan Liu, Haoming Jiang, Pengcheng He, Weizhu Chen et al.ICLR 2020 · 2,210 citations
- Natural Image Matting via Guided Contextual AttentionYaoyi Li, Hongtao LuAAAI 2020 · 189 citations
- MatteFormer: Transformer-Based Image Matting via Prior-TokensGyutae Park, Sungjoon Son, Jaeyoung Yoo, Seho Kim et al.CVPR 2022 · 82 citations
- Tripartite Information Mining and Integration for Image MattingYuhao Liu, Jiake Xie, Xiao Shi, Yu Qiao et al.ICCV 2021 · 66 citations
Related papers
- Context-Aware Image Matting for Simultaneous Foreground and Alpha EstimationQiqi Hou, Feng LiuICCV 2019 · 171 citations
- High-Resolution Deep Image MattingHaichao Yu, Ning Xu, Zilong Huang, Yuqian Zhou et al.AAAI 2021 · 61 citations
- Attention-Guided Hierarchical Structure Aggregation for Image MattingYu Qiao, Yuhao Liu, Xin Yang, Dongsheng Zhou et al.CVPR 2020
- Indices Matter: Learning to Index for Deep Image MattingHao Lu, Yutong Dai, Chunhua Shen, Songcen XuICCV 2019 · 206 citations
- Deep Video Matting via Spatio-Temporal Alignment and AggregationYanan Sun, Guanzhi Wang, Qiao Gu, Chi-Keung Tang et al.CVPR 2021
